paper-with-me

Papers

OCM3D: Object-Centric Monocular 3D Object Detection

2021-04-13 · Liang Peng, Fei Liu, Senbo Yan, Xiaofei He, Deng Cai

Image-only and pseudo-LiDAR representations are commonly used for monocular 3D object detection. However, methods based on them have shortcomings of either not well capturing the spatial relationships in neighbored image pixels or being hard to handle the noisy nature of the monocular pseudo-LiDAR point cloud. To overcome these issues, in this paper we propose a novel object-centric voxel representation tailored for monocular 3D object detection. Specifically, voxels are built on each object proposal, and their sizes are adaptively determined by the 3D spatial distribution of the points, allowing the noisy point cloud to be organized effectively within a voxel grid. This representation is proved to be able to locate the object in 3D space accurately. Furthermore, prior works would like to estimate the orientation via deep features extracted from an entire image or a noisy point cloud. By contrast, we argue that the local RoI information from the object image patch alone with a proper resizing scheme is a better input as it provides complete semantic clues meanwhile excludes irrelevant interferences. Besides, we decompose the confidence mechanism in monocular 3D object detection by considering the relationship between 3D objects and the associated 2D boxes. Evaluated on KITTI, our method outperforms state-of-the-art methods by a large margin. The code will be made publicly available soon.

📄 PDF Abstract BibTeX arXiv:2104.06041

Code (0)

등록된 구현이 없습니다.

Tasks

3D Object DetectionMonocular 3D Object DetectionObjectobject-detectionObject Detection

Similar Papers 제목 키워드 기반

Boosting Monocular 3D Object Detection with Object-Centric Auxiliary Depth Supervision

2022-10-29 · Youngseok Kim, Sanmin Kim, Sangmin Sim, Jun Won Choi 외

Recent advances in monocular 3D detection leverage a depth estimation network explicitly as an intermediate stage of the 3D detection network. Depth map approaches yield more accurate depth to objects than other methods …

3D Object DetectionDepth EstimationDepth PredictionMonocular 3D Object Detection+4

Learning Geocentric Object Pose in Oblique Monocular Images

2020-07-01 · CVPR 2020 6 · Gordon Christie, Rodrigo Rene Rai Munoz Abujder, Kevin Foster, Shea Hagstrom 외

An object's geocentric pose, defined as the height above ground and orientation with respect to gravity, is a powerful representation of real-world structure for object detection, segmentation, and localization tasks usi…

Earth ObservationObjectobject-detectionObject Detection+2

MOLTR: Multiple Object Localisation, Tracking, and Reconstruction from Monocular RGB Videos

2020-12-09 · Kejie Li, Hamid Rezatofighi, Ian Reid

Semantic aware reconstruction is more advantageous than geometric-only reconstruction for future robotic and AR/VR applications because it represents not only where things are, but also what things are. Object-centric ma…

BenchmarkingObjectObject Localization

3D Hand Pose Detection in Egocentric RGB-D Images

2014-11-29 · Gregory Rogez, James S. Supancic III, Maryam Khademi, Jose Maria Martinez Montiel 외

We focus on the task of everyday hand pose estimation from egocentric viewpoints. For this task, we show that depth sensors are particularly informative for extracting near-field interactions of the camera wearer with hi…

Hand DetectionHand Pose EstimationPose Estimation

RaGS: Unleashing 3D Gaussian Splatting from 4D Radar and Monocular Cues for 3D Object Detection

2025-07-26 · Xiaokai Bai, Chenxu Zhou, Lianqing Zheng, Si-Yuan Cao 외 arxiv

4D millimeter-wave radar is a promising sensing modality for autonomous driving, yet effective 3D object detection from 4D radar and monocular images remains challenging. Existing fusion approaches either rely on instanc…

3D Object DetectionAutonomous Driving