IDA-3D: Instance-Depth-Aware 3D Object Detection From Stereo Vision for Autonomous Driving
3D object detection is an important scene understanding task in autonomous driving and virtual reality. Approaches based on LiDAR technology have high performance, but LiDAR is expensive. Considering more general scenes, where there is no LiDAR data in the 3D datasets, we propose a 3D object detection approach from stereo vision which does not rely on LiDAR data either as input or as supervision in training, but solely takes RGB images with corresponding annotated 3D bounding boxes as training data. As depth estimation of object is the key factor affecting the performance of 3D object detection, we introduce an Instance-DepthAware (IDA) module which accurately predicts the depth of the 3D bounding box's center by instance-depth awareness, disparity adaptation and matching cost reweighting. Moreover, our model is an end-to-end learning framework which does not require multiple stages or postprocessing algorithm. We provide detailed experiments on KITTI benchmark and achieve impressive improvements compared with the existing image-based methods. Our code is available at https://github.com/swords123/IDA-3D.
Code (1)
Tasks
3D Object DetectionAutonomous DrivingDepth EstimationObjectobject-detectionObject DetectionScene UnderstandingSimilar Papers 제목 키워드 기반
SIDE: Center-based Stereo 3D Detector with Structure-aware Instance Depth Estimation
3D detection plays an indispensable role in environment perception. Due to the high cost of commonly used LiDAR sensor, stereo vision based 3D detection, as an economical yet effective setting, attracts more attention re…
Depth EstimationDepth-Aware Rover: A Study of Edge AI and Monocular Vision for Real-World Implementation
This study analyses simulated and real-world implementations of depth-aware rover navigation, highlighting the transition from stereo vision to monocular depth estimation using edge AI. A Unity-based lunar terrain simula…
Monocular Depth EstimationReal-Time Object DetectionUnknown Object Segmentation from Stereo Images
Although instance-aware perception is a key prerequisite for many autonomous robotic applications, most of the methods only partially solve the problem by focusing solely on known object categories. However, for robots i…
Instance SegmentationObjectSegmentationSemantic SegmentationInstance-aware Multi-Camera 3D Object Detection with Structural Priors Mining and Self-Boosting Learning
Camera-based bird-eye-view (BEV) perception paradigm has made significant progress in the autonomous driving field. Under such a paradigm, accurate BEV representation construction relies on reliable depth estimation for …
3D Object DetectionAutonomous DrivingDepth Estimationobject-detection+2LIGA-Stereo: Learning LiDAR Geometry Aware Representations for Stereo-based 3D Detector
Stereo-based 3D detection aims at detecting 3D object bounding boxes from stereo images using intermediate depth maps or implicit 3D geometry representations, which provides a low-cost solution for 3D perception. However…
3D geometry3D Object Detection From Stereo ImagesStereo Matching