Pix2Point: Learning Outdoor 3D Using Sparse Point Clouds and Optimal Transport
Good quality reconstruction and comprehension of a scene rely on 3D estimation methods. The 3D information was usually obtained from images by stereo-photogrammetry, but deep learning has recently provided us with excellent results for monocular depth estimation. Building up a sufficiently large and rich training dataset to achieve these results requires onerous processing. In this paper, we address the problem of learning outdoor 3D point cloud from monocular data using a sparse ground-truth dataset. We propose Pix2Point, a deep learning-based approach for monocular 3D point cloud prediction, able to deal with complete and challenging outdoor scenes. Our method relies on a 2D-3D hybrid neural network architecture, and a supervised end-to-end minimisation of an optimal transport divergence between point clouds. We show that, when trained on sparse point clouds, our simple promising approach achieves a better coverage of 3D outdoor scenes than efficient monocular depth methods.
Code (0)
등록된 구현이 없습니다.
Tasks
Deep LearningDepth EstimationMonocular Depth EstimationSimilar Papers 제목 키워드 기반
MNEW: Multi-domain Neighborhood Embedding and Weighting for Sparse Point Clouds Segmentation
Point clouds have been widely adopted in 3D semantic scene understanding. However, point clouds for typical tasks such as 3D shape segmentation or indoor scenario parsing are much denser than outdoor LiDAR sweeps for the…
Autonomous DrivingScene UnderstandingSemantic SegmentationMulti-modal panoramic 3D outdoor datasets for place categorization
We present two multi-modal panoramic 3D outdoor (MPO) datasets for semantic place categorization with six categories: forest, coast, residential area, urban area and indoor/outdoor parking lot. The first dataset consists…
Point CloudsDiffusion-Based Point Cloud Super-Resolution for mmWave Radar Data
The millimeter-wave radar sensor maintains stable performance under adverse environmental conditions, making it a promising solution for all-weather perception tasks, such as outdoor mobile robotics. However, the radar p…
Point Cloud Super ResolutionSuper-ResolutionOccupancy-MAE: Self-supervised Pre-training Large-scale LiDAR Point Clouds with Masked Occupancy Autoencoders
Current perception models in autonomous driving heavily rely on large-scale labelled 3D data, which is both costly and time-consuming to annotate. This work proposes a solution to reduce the dependence on labelled 3D tra…
3D Object Detection3D Semantic SegmentationAutonomous DrivingDomain Adaptation+8LidaRefer: Outdoor 3D Visual Grounding for Autonomous Driving with Transformers
3D visual grounding (VG) aims to locate relevant objects or regions within 3D scenes based on natural language descriptions. Although recent methods for indoor 3D VG have successfully transformer-based architectures to c…
3D visual groundingAutonomous DrivingVisual Grounding