Object Localization under Single Coarse Point Supervision
Point-based object localization (POL), which pursues high-performance object sensing under low-cost data annotation, has attracted increased attention. However, the point annotation mode inevitably introduces semantic variance for the inconsistency of annotated points. Existing POL methods heavily reply on accurate key-point annotations which are difficult to define. In this study, we propose a POL method using coarse point annotations, relaxing the supervision signals from accurate key points to freely spotted points. To this end, we propose a coarse point refinement (CPR) approach, which to our best knowledge is the first attempt to alleviate semantic variance from the perspective of algorithm. CPR constructs point bags, selects semantic-correlated points, and produces semantic center points through multiple instance learning (MIL). In this way, CPR defines a weakly supervised evolution procedure, which ensures training high-performance object localizer under coarse point supervision. Experimental results on COCO, DOTA and our proposed SeaPerson dataset validate the effectiveness of the CPR approach. The dataset and code will be available at https://github.com/ucas-vg/PointTinyBenchmark/.
Code (2)
Tasks
Multiple Instance LearningObjectObject LocalizationSimilar Papers 제목 키워드 기반
CPR++: Object Localization via Single Coarse Point Supervision
Point-based object localization (POL), which pursues high-performance object sensing under low-cost data annotation, has attracted increased attention. However, the point annotation mode inevitably introduces semantic va…
ObjectObject LocalizationP2P-Loc: Point to Point Tiny Person Localization
Bounding-box annotation form has been the most frequently used method for visual object localization tasks. However, bounding-box annotation relies on a large amount of precisely annotating bounding boxes, and it is expe…
ObjectObject LocalizationRepPoints: Point Set Representation for Object Detection
Modern object detectors rely heavily on rectangular bounding boxes, such as anchors, proposals and the final predictions, to represent objects at various recognition stages. The bounding box is convenient to use but prov…
Objectobject-detectionObject DetectionG2L-Net: Global to Local Network for Real-time 6D Pose Estimation with Embedding Vector Features
In this paper, we propose a novel real-time 6D object pose estimation framework, named G2L-Net. Our network operates on point clouds from RGB-D detection in a divide-and-conquer fashion. Specifically, our network consist…
6D Pose Estimation6D Pose Estimation using RGBObjectPose Estimation+1Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization
Weakly Supervised Object Localization (WSOL), which aims to localize objects by only using image-level labels, has attracted much attention because of its low annotation cost in real applications. Current studies focus o…
Object LocalizationWeakly-Supervised Object LocalizationZero-shot Generalization