paper-with-me

Papers

MoDAR: Using Motion Forecasting for 3D Object Detection in Point Cloud Sequences

2023-06-05 · CVPR 2023 1 · Yingwei Li, Charles R. Qi, Yin Zhou, Chenxi Liu, Dragomir Anguelov

Occluded and long-range objects are ubiquitous and challenging for 3D object detection. Point cloud sequence data provide unique opportunities to improve such cases, as an occluded or distant object can be observed from different viewpoints or gets better visibility over time. However, the efficiency and effectiveness in encoding long-term sequence data can still be improved. In this work, we propose MoDAR, using motion forecasting outputs as a type of virtual modality, to augment LiDAR point clouds. The MoDAR modality propagates object information from temporal contexts to a target frame, represented as a set of virtual points, one for each object from a waypoint on a forecasted trajectory. A fused point cloud of both raw sensor points and the virtual points can then be fed to any off-the-shelf point-cloud based 3D object detector. Evaluated on the Waymo Open Dataset, our method significantly improves prior art detectors by using motion forecasting from extra-long sequences (e.g. 18 seconds), achieving new state of the arts, while not adding much computation overhead.

📄 PDF Abstract BibTeX arXiv:2306.03206

Code (1)

quan-dao/practical-collab-perception pytorch

Tasks

3D Object DetectionMotion ForecastingObjectobject-detectionObject Detection

Similar Papers 제목 키워드 기반

Modality-Autoregressive World-Action Models

2026-09-15 · Adam Hung, Bardienus P. Duisterhof, Deva Ramanan, Jeffrey Ichnowski arxiv

World-action models (WAMs) jointly model future observations and actions, typically predicting the future as RGB images. Other visual modalities such as depth, pretrained visual features, and point tracks can more effici…

emoDARTS: Joint Optimisation of CNN & Sequential Neural Network Architectures for Superior Speech Emotion Recognition

2024-03-21 · Thejan Rajapakshe, Rajib Rana, Sara Khalifa, Berrak Sisman 외

Speech Emotion Recognition (SER) is crucial for enabling computers to understand the emotions conveyed in human communication. With recent advancements in Deep Learning (DL), the performance of SER models has significant…

Emotion RecognitionNeural Architecture SearchSpeech Emotion Recognition

MolmoMotion: Forecasting Point Trajectories in 3D with Language Instruction

2026-06-17 · Jianing Zhang, Chenhao Zheng, Yajun Yang, Max Argus 외 arxiv

Motion forecasting is central to visual intelligence: agents must anticipate how objects will move in order to plan actions, reason about physical interactions, and synthesize realistic futures. We argue that 3D points i…

Robot ManipulationMotion Forecasting

RV-FuseNet: Range View Based Fusion of Time-Series LiDAR Data for Joint 3D Object Detection and Motion Forecasting

2020-05-21 · Ankit Laddha, Shivam Gautam, Gregory P. Meyer, Carlos Vallespi-Gonzalez 외

Robust real-time detection and motion forecasting of traffic participants is necessary for autonomous vehicles to safely navigate urban environments. In this paper, we present RV-FuseNet, a novel end-to-end approach for …

3D Object DetectionAutonomous DrivingAutonomous VehiclesMotion Forecasting+6

Future Does Matter: Boosting 3D Object Detection with Temporal Motion Estimation in Point Cloud Sequences

2024-09-06 · Rui Yu, Runkai Zhao, Cong Nie, Heng Wang 외

Accurate and robust LiDAR 3D object detection is essential for comprehensive scene understanding in autonomous driving. Despite its importance, LiDAR detection performance is limited by inherent constraints of point clou…

3D Object DetectionAutonomous DrivingMotion EstimationMotion Forecasting+4