paper-with-me

Papers

Lifting Object Detection Datasets into 3D

2015-03-22 · Joao Carreira, Sara Vicente, Lourdes Agapito, Jorge Batista

While data has certainly taken the center stage in computer vision in recent years, it can still be difficult to obtain in certain scenarios. In particular, acquiring ground truth 3D shapes of objects pictured in 2D images remains a challenging feat and this has hampered progress in recognition-based object reconstruction from a single image. Here we propose to bypass previous solutions such as 3D scanning or manual design, that scale poorly, and instead populate object category detection datasets semi-automatically with dense, per-object 3D reconstructions, bootstrapped from:(i) class labels, (ii) ground truth figure-ground segmentations and (iii) a small set of keypoint annotations. Our proposed algorithm first estimates camera viewpoint using rigid structure-from-motion and then reconstructs object shapes by optimizing over visual hull proposals guided by loose within-class shape similarity assumptions. The visual hull sampling process attempts to intersect an object's projection cone with the cones of minimal subsets of other similar objects among those pictured from certain vantage points. We show that our method is able to produce convincing per-object 3D reconstructions and to accurately estimate cameras viewpoints on one of the most challenging existing object-category detection datasets, PASCAL VOC. We hope that our results will re-stimulate interest on joint object recognition and 3D reconstruction from a single image.

📄 PDF Abstract BibTeX arXiv:1503.06465

Code (0)

등록된 구현이 없습니다.

Tasks

3D ReconstructionObjectobject-detectionObject DetectionObject RecognitionObject Reconstruction

Similar Papers 제목 키워드 기반

ROI-10D: Monocular Lifting of 2D Detection to 6D Pose and Metric Shape

2018-12-06 · CVPR 2019 6 · Fabian Manhardt, Wadim Kehl, Adrien Gaidon

We present a deep learning method for end-to-end monocular 3D object detection and metric shape retrieval. We propose a novel loss formulation by lifting 2D detection, orientation, and scale estimation into 3D space. Ins…

3D Object DetectionData AugmentationMonocular 3D Object Detectionobject-detection+2

Boxer: Robust Lifting of Open-World 2D Bounding Boxes to 3D

2026-04-06 · Daniel DeTone, Tianwei Shen, Fan Zhang, Lingni Ma 외 arxiv

Detecting and localizing objects in space is a fundamental computer vision problem. While much progress has been made to solve 2D object detection, 3D object localization is much less explored and far from solved, especi…

Object Localization2D Object Detection

Multi-View Attentive Contextualization for Multi-View 3D Object Detection

2024-05-20 · CVPR 2024 1 · Xianpeng Liu, Ce Zheng, Ming Qian, Nan Xue 외

We present Multi-View Attentive Contextualization (MvACon), a simple yet effective method for improving 2D-to-3D feature lifting in query-based multi-view 3D (MV3D) object detection. Despite remarkable progress witnessed…

3D Object DetectionObjectobject-detectionObject Detection

Lifting Multi-View Detection and Tracking to the Bird's Eye View

2024-03-19 · Torben Teepe, Philipp Wolters, Johannes Gilg, Fabian Herzog 외

Taking advantage of multi-view aggregation presents a promising solution to tackle challenges such as occlusion and missed detection in multi-object tracking and detection. Recent advancements in multi-view detection and…

3D Object RecognitionMulti-Object Trackingmulti-view detectionMultiview Detection+2

Learning-based safety lifting monitoring system for cranes on construction sites

2025-06-25 · Hao Chen, Yu Hin Ng, Ching-Wei Chang, Haobo Liang 외

Lifting on construction sites, as a frequent operation, works still with safety risks, especially for modular integrated construction (MiC) lifting due to its large weight and size, probably leading to accidents, causing…