Simultaneous multi-view instance detection with learned geometric soft-constraints
We propose to jointly learn multi-view geometry and warping between views of the same object instances for robust cross-view object detection. What makes multi-view object instance detection difficult are strong changes in viewpoint, lighting conditions, high similarity of neighbouring objects, and strong variability in scale. By turning object detection and instance re-identification in different views into a joint learning task, we are able to incorporate both image appearance and geometric soft constraints into a single, multi-view detection process that is learnable end-to-end. We validate our method on a new, large data set of street-level panoramas of urban objects and show superior performance compared to various baselines. Our contribution is threefold: a large-scale, publicly available data set for multi-view instance detection and re-identification; an annotation tool custom-tailored for multi-view instance detection; and a novel, holistic multi-view instance detection and re-identification method that jointly models geometry and appearance across views.
Code (0)
등록된 구현이 없습니다.
Tasks
multi-view detectionObjectobject-detectionObject DetectionSimilar Papers 제목 키워드 기반
A Unified Query-based Paradigm for Camouflaged Instance Segmentation
Due to the high similarity between camouflaged instances and the background, the recently proposed camouflaged instance segmentation (CIS) faces challenges in accurate localization and instance segmentation. To this end,…
Boundary DetectionDecoderInstance SegmentationMulti-Task Learning+2Consistent Instance Classification for Unsupervised Representation Learning
In this paper, we address the problem of learning the representations from images without human annotations. We study the instance classification solution, which regards each instance as a category, and improve the optim…
ClassificationGeneral ClassificationLinear evaluationRepresentation LearningMooMIns -- Monocular 3D Reconstruction and Object Pose Estimation from Multiple Instances
Simultaneous 3D reconstruction and 6D object pose estimation from a single monocular image is an inherently ill-posed problem. In industrial settings, however, multiple instances of an object are often randomly arranged …
Monocular Depth EstimationInstance Segmentation3D ReconstructionPose EstimationUniversal Representation Learning of Knowledge Bases by Jointly Embedding Instances and Ontological Concepts
Many large-scale knowledge bases simultaneously represent two views of knowledge graphs (KGs): an ontology view for abstract and commonsense concepts, and an instance view for specific entities that are instantiated from…
Entity TypingKnowledge GraphsRepresentation LearningTowards View-invariant and Accurate Loop Detection Based on Scene Graph
Loop detection plays a key role in visual Simultaneous Localization and Mapping (SLAM) by correcting the accumulated pose drift. In indoor scenarios, the richly distributed semantic landmarks are view-point invariant and…
DescriptiveSimultaneous Localization and Mapping