paper-with-me

Papers

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving

2025-04-17 · Shumin Wang, Zhuoran Yang, Lidian Wang, Zhipeng Tang, Heng Li, Lehan Pan, Sha Zhang, Jie Peng, Jianmin Ji, Yanyong Zhang

The significant achievements of pre-trained models leveraging large volumes of data in the field of NLP and 2D vision inspire us to explore the potential of extensive data pre-training for 3D perception in autonomous driving. Toward this goal, this paper proposes to utilize massive unlabeled data from heterogeneous datasets to pre-train 3D perception models. We introduce a self-supervised pre-training framework that learns effective 3D representations from scratch on unlabeled data, combined with a prompt adapter based domain adaptation strategy to reduce dataset bias. The approach significantly improves model performance on downstream tasks such as 3D object detection, BEV segmentation, 3D object tracking, and occupancy prediction, and shows steady performance increase as the training data volume scales up, demonstrating the potential of continually benefit 3D perception models for autonomous driving. We will release the source code to inspire further investigations in the community.

📄 PDF Abstract BibTeX arXiv:2504.12709

Code (0)

등록된 구현이 없습니다.

Tasks

3D Object Detection3D Object TrackingAutonomous DrivingBEV SegmentationDomain Adaptationobject-detectionObject DetectionObject Tracking

Methods 이 논문이 사용한 방법론

Adapter 설명 없음

Similar Papers 제목 키워드 기반

SpikeCLR: Contrastive Self-Supervised Learning for Few-Shot Event-Based Vision using Spiking Neural Networks

2026-03-17 · Maxime Vaillant, Axel Carlier, Lai Xing Ng, Christophe Hurter 외 arxiv

Event-based vision sensors provide significant advantages for high-speed perception, including microsecond temporal resolution, high dynamic range, and low power consumption. When combined with Spiking Neural Networks (S…

Self-Supervised LearningEvent-based vision

Learning Shared RGB-D Fields: Unified Self-supervised Pre-training for Label-efficient LiDAR-Camera 3D Perception

2024-05-28 · Xiaohao Xu, Ye Li, Tianyi Zhang, Jinrong Yang 외

Constructing large-scale labeled datasets for multi-modal perception model training in autonomous driving presents significant challenges. This has motivated the development of self-supervised pretraining strategies. How…

3D Object DetectionAutonomous DrivingNeRFNeural Rendering+4

Towards Unsupervised Object Detection From LiDAR Point Clouds

2023-11-03 · CVPR 2023 1 · Lunjun Zhang, Anqi Joyce Yang, Yuwen Xiong, Sergio Casas 외

In this paper, we study the problem of unsupervised object detection from 3D point clouds in self-driving scenes. We present a simple yet effective method that exploits (i) point clustering in near-range areas where the …

Objectobject-detectionObject DetectionObject Discovery+1

Boosting Supervision with Self-Supervision for Few-shot Learning

2019-06-17 · Jong-Chyi Su, Subhransu Maji, Bharath Hariharan

We present a technique to improve the transferability of deep representations learned on small labeled datasets by introducing self-supervised tasks as auxiliary loss functions. While recent approaches for self-supervise…

Few-Shot LearningSelf-Supervised Learning

Self-supervised Learning for Sonar Image Classification

2022-04-20 · Alan Preciado-Grijalva, Bilal Wehbe, Miguel Bande Firvida, Matias Valdenegro-Toro

Self-supervised learning has proved to be a powerful approach to learn image representations without the need of large labeled datasets. For underwater robotics, it is of great interest to design computer vision algorith…

ClassificationDenoisingimage-classificationImage Classification+2