Self-Supervised Pillar Motion Learning for Autonomous Driving
Autonomous driving can benefit from motion behavior comprehension when interacting with diverse traffic participants in highly dynamic environments. Recently, there has been a growing interest in estimating class-agnostic motion directly from point clouds. Current motion estimation methods usually require vast amount of annotated training data from self-driving scenes. However, manually labeling point clouds is notoriously difficult, error-prone and time-consuming. In this paper, we seek to answer the research question of whether the abundant unlabeled data collections can be utilized for accurate and efficient motion learning. To this end, we propose a learning framework that leverages free supervisory signals from point clouds and paired camera images to estimate motion purely via self-supervision. Our model involves a point cloud based structural consistency augmented with probabilistic motion masking as well as a cross-sensor motion regularization to realize the desired self-supervision. Experiments reveal that our approach performs competitively to supervised methods, and achieves the state-of-the-art result when combining our self-supervised model with supervised fine-tuning.
Code (1)
Tasks
Autonomous DrivingMotion EstimationSimilar Papers 제목 키워드 기반
ContrastMotion: Self-supervised Scene Motion Learning for Large-Scale LiDAR Point Clouds
In this paper, we propose a novel self-supervised motion estimator for LiDAR-based autonomous driving via BEV representation. Different from usually adopted self-supervised strategies for data-level structure consistency…
Autonomous DrivingContrastive Learningmotion predictionvalidVoteFlow: Enforcing Local Rigidity in Self-Supervised Scene Flow
Scene flow estimation aims to recover per-point motion from two adjacent LiDAR scans. However, in real-world applications such as autonomous driving, points rarely move independently of others, especially for nearby poin…
Autonomous DrivingComputational EfficiencyInductive BiasScene Flow Estimation+1Mapillary Vistas Validation for Fine-Grained Traffic Signs: A Benchmark Revealing Vision-Language Model Limitations
Obtaining high-quality fine-grained annotations for traffic signs is critical for accurate and safe decision-making in autonomous driving. Widely used datasets, such as Mapillary, often provide only coarse-grained labels…
Traffic Sign RecognitionAutonomous DrivingJust Go with the Flow: Self-Supervised Scene Flow Estimation
When interacting with highly dynamic environments, scene flow allows autonomous systems to reason about the non-rigid motion of multiple independent objects. This is of particular interest in the field of autonomous driv…
Autonomous DrivingScene Flow EstimationSelf-supervised Scene Flow EstimationSSC3OD: Sparsely Supervised Collaborative 3D Object Detection from LiDAR Point Clouds
Collaborative 3D object detection, with its improved interaction advantage among multiple agents, has been widely explored in autonomous driving. However, existing collaborative 3D object detectors in a fully supervised …
3D Object DetectionAutonomous DrivingObjectobject-detection+1