View-to-Label: Multi-View Consistency for Self-Supervised 3D Object Detection
For autonomous vehicles, driving safely is highly dependent on the capability to correctly perceive the environment in 3D space, hence the task of 3D object detection represents a fundamental aspect of perception. While 3D sensors deliver accurate metric perception, monocular approaches enjoy cost and availability advantages that are valuable in a wide range of applications. Unfortunately, training monocular methods requires a vast amount of annotated data. Interestingly, self-supervised approaches have recently been successfully applied to ease the training process and unlock access to widely available unlabelled data. While related research leverages different priors including LIDAR scans and stereo images, such priors again limit usability. Therefore, in this work, we propose a novel approach to self-supervise 3D object detection purely from RGB sequences alone, leveraging multi-view constraints and weak labels. Our experiments on KITTI 3D dataset demonstrate performance on par with state-of-the-art self-supervised methods using LIDAR scans or stereo images.
Code (0)
등록된 구현이 없습니다.
Tasks
3D Object DetectionAutonomous Vehiclesobject-detectionObject DetectionSimilar Papers 제목 키워드 기반
EICO: Improving Few-Shot Text Classification via Explicit and Implicit Consistency Regularization
While the prompt-based fine-tuning methods had advanced few-shot natural language understanding tasks, self-training methods are also being explored. This work revisits the consistency regularization in self-training and…
Few-Shot LearningFew-Shot Text ClassificationLanguage ModelingLanguage Modelling+4HaMuCo: Hand Pose Estimation via Multiview Collaborative Self-Supervised Learning
Recent advancements in 3D hand pose estimation have shown promising results, but its effectiveness has primarily relied on the availability of large-scale annotated datasets, the creation of which is a laborious and cost…
3D Hand Pose EstimationHand Pose EstimationPose EstimationSelf-Supervised Learning360-MLC: Multi-view Layout Consistency for Self-training and Hyper-parameter Tuning
We present 360-MLC, a self-training method based on multi-view layout consistency for finetuning monocular room-layout models using unlabeled 360-images only. This can be valuable in practical scenarios where a pre-train…
Model SelectionPseudo LabelSelf-Supervised Cross-View Correspondence with Predictive Cycle Consistency
Learning self-supervised visual correspondence is a long-studied task fundamental to visual understanding and human perception. However, existing correspondence methods largely focus on small image transformations, s…
ColorizationImitation LearningObject TrackingMulti-View Correlation Consistency for Semi-Supervised Semantic Segmentation
Semi-supervised semantic segmentation needs rich and robust supervision on unlabeled data. Consistency learning enforces the same pixel to have similar features in different augmented views, which is a robust signal but …
Contrastive LearningData AugmentationSemantic SegmentationSemi-Supervised Semantic Segmentation