Aug3D-RPN: Improving Monocular 3D Object Detection by Synthetic Images with Virtual Depth
Current geometry-based monocular 3D object detection models can efficiently detect objects by leveraging perspective geometry, but their performance is limited due to the absence of accurate depth information. Though this issue can be alleviated in a depth-based model where a depth estimation module is plugged to predict depth information before 3D box reasoning, the introduction of such module dramatically reduces the detection speed. Instead of training a costly depth estimator, we propose a rendering module to augment the training data by synthesizing images with virtual-depths. The rendering module takes as input the RGB image and its corresponding sparse depth image, outputs a variety of photo-realistic synthetic images, from which the detection model can learn more discriminative features to adapt to the depth changes of the objects. Besides, we introduce an auxiliary module to improve the detection model by jointly optimizing it through a depth estimation task. Both modules are working in the training time and no extra computation will be introduced to the detection model. Experiments show that by working with our proposed modules, a geometry-based model can represent the leading accuracy on the KITTI 3D detection benchmark.
Code (0)
등록된 구현이 없습니다.
Tasks
3D Object DetectionDepth EstimationMonocular 3D Object Detectionobject-detectionObject DetectionSimilar Papers 제목 키워드 기반
CubifAE-3D: Monocular Camera Space Cubification for Auto-Encoder based 3D Object Detection
We introduce a method for 3D object detection using a single monocular image. Starting from a synthetic dataset, we pre-train an RGB-to-Depth Auto-Encoder (AE). The embedding learnt from this AE is then used to train a 3…
3D Object Detection3D Object Detection From Monocular ImagesAutonomous VehiclesMonocular 3D Object Detection+3Virtually Enriched NYU Depth V2 Dataset for Monocular Depth Estimation: Do We Need Artificial Augmentation?
We present ANYU, a new virtually augmented version of the NYU depth v2 dataset, designed for monocular depth estimation. In contrast to the well-known approach where full 3D scenes of a virtual world are utilized to gene…
Depth EstimationMonocular Depth Estimation3D Object Detection from a Single Fisheye Image Without a Single Fisheye Training Image
Existing monocular 3D object detection methods have been demonstrated on rectilinear perspective images and fail in images with alternative projections such as those acquired by fisheye cameras. Previous works on object …
2D Object Detection3D Object DetectionMonocular 3D Object DetectionObject+23D Copy-Paste: Physically Plausible Object Insertion for Monocular 3D Detection
A major challenge in monocular 3D object detection is the limited diversity and quantity of objects in real datasets. While augmenting real scenes with virtual objects holds promise to improve both the diversity and quan…
3D Object DetectionData AugmentationDiversityMonocular 3D Object Detection+3Towards Generalization Across Depth for Monocular 3D Object Detection
While expensive LiDAR and stereo camera rigs have enabled the development of successful 3D object detection methods, monocular RGB-only approaches lag much behind. This work advances the state of the art by introducing M…
3D Object DetectionMonocular 3D Object DetectionObjectobject-detection+1