paper-with-me

Papers

Learning to Detect Objects from Multi-Agent LiDAR Scans without Manual Labels

2025-03-11 · CVPR 2025 1 · Qiming Xia, Wenkai Lin, Haoen Xiang, Xun Huang, Siheng Chen, Zhen Dong, Cheng Wang, Chenglu Wen

Unsupervised 3D object detection serves as an important solution for offline 3D object annotation. However, due to the data sparsity and limited views, the clustering-based label fitting in unsupervised object detection often generates low-quality pseudo-labels. Multi-agent collaborative dataset, which involves the sharing of complementary observations among agents, holds the potential to break through this bottleneck. In this paper, we introduce a novel unsupervised method that learns to Detect Objects from Multi-Agent LiDAR scans, termed DOtA, without using labels from external. DOtA first uses the internally shared ego-pose and ego-shape of collaborative agents to initialize the detector, leveraging the generalization performance of neural networks to infer preliminary labels. Subsequently,DOtA uses the complementary observations between agents to perform multi-scale encoding on preliminary labels, then decodes high-quality and low-quality labels. These labels are further used as prompts to guide a correct feature learning process, thereby enhancing the performance of the unsupervised object detection task. Extensive experiments on the V2V4Real and OPV2V datasets show that our DOtA outperforms state-of-the-art unsupervised 3D object detection methods. Additionally, we also validate the effectiveness of the DOtA labels under various collaborative perception frameworks.The code is available at https://github.com/xmuqimingxia/DOtA.

📄 PDF Abstract BibTeX arXiv:2503.08421

Code (1)

xmuqimingxia/dota 공식 구현 pytorch

Tasks

3D Object DetectionObjectobject-detectionObject DetectionUnsupervised Object Detection

Similar Papers 제목 키워드 기반

TwistSLAM++: Fusing multiple modalities for accurate dynamic semantic SLAM

2022-09-16 · Mathieu Gonzalez, Eric Marchand, Amine Kacete, Jérôme Royan

Most classical SLAM systems rely on the static scene assumption, which limits their applicability in real world scenarios. Recent SLAM frameworks have been proposed to simultaneously track the camera and moving objects. …

ObjectObject TrackingPose EstimationSemantic SLAM

PanoNet3D: Combining Semantic and Geometric Understanding for LiDARPoint Cloud Detection

2020-12-17 · Xia Chen, Jianren Wang, David Held, Martial Hebert

Visual data in autonomous driving perception, such as camera image and LiDAR point cloud, can be interpreted as a mixture of two aspects: semantic feature and geometric structure. Semantics come from the appearance and c…

Autonomous DrivingCloud Detection

Interactive4D: Interactive 4D LiDAR Segmentation

2024-10-10 · Ilya Fradlin, Idil Esen Zulfikar, Kadir Yilmaz, Theodora Kontogianni 외

Interactive segmentation has an important role in facilitating the annotation process of future LiDAR datasets. Existing approaches sequentially segment individual objects at each LiDAR scan, repeating the process throug…

Interactive SegmentationSegmentation

VANETs Meet Autonomous Vehicles: A Multimodal 3D Environment Learning Approach

2017-05-24 · Yassine Maalej, Sameh Sorour, Ahmed Abdel-Rahim, Mohsen Guizani

In this paper, we design a multimodal framework for object detection, recognition and mapping based on the fusion of stereo camera frames, point cloud Velodyne Lidar scans, and Vehicle-to-Vehicle (V2V) Basic Safety Messa…

Autonomous Vehiclesobject-detectionObject Detection

MOVES: Movable and Moving LiDAR Scene Segmentation in Label-Free settings using Static Reconstruction

2023-06-26 · Prashant Kumar, Dhruv Makwana, Onkar Susladkar, Anurag Mittal 외

Accurate static structure reconstruction and segmentation of non-stationary objects is of vital importance for autonomous navigation applications. These applications assume a LiDAR scan to consist of only static structur…

Autonomous NavigationScene SegmentationSegmentation