paper-with-me

홈 › Papers

M3CAD: Towards Generic Cooperative Autonomous Driving Benchmark

2025-05-10 · Morui Zhu, Yongqi Zhu, Yihao Zhu, Qi Chen, Deyuan Qu, Song Fu, Qing Yang

We introduce M$^3$CAD, a novel benchmark designed to advance research in generic cooperative autonomous driving. M$^3$CAD comprises 204 sequences with 30k frames, spanning a diverse range of cooperative driving scenarios. Each sequence includes multiple vehicles and sensing modalities, e.g., LiDAR point clouds, RGB images, and GPS/IMU, supporting a variety of autonomous driving tasks, including object detection and tracking, mapping, motion forecasting, occupancy prediction, and path planning. This rich multimodal setup enables M$^3$CAD to support both single-vehicle and multi-vehicle autonomous driving research, significantly broadening the scope of research in the field. To our knowledge, M$^3$CAD is the most comprehensive benchmark specifically tailored for cooperative multi-task autonomous driving research. We evaluate the state-of-the-art end-to-end solution on M$^3$CAD to establish baseline performance. To foster cooperative autonomous driving research, we also propose E2EC, a simple yet effective framework for cooperative driving solution that leverages inter-vehicle shared information for improved path planning. We release M$^3$CAD, along with our baseline models and evaluation results, to support the development of robust cooperative autonomous driving systems. All resources will be made publicly available on https://github.com/zhumorui/M3CAD

📄 PDF Abstract BibTeX arXiv:2505.06746

Code (1)

zhumorui/m3cad 공식 구현 pytorch

Tasks

Autonomous DrivingMotion Forecastingobject-detectionObject Detection

Similar Papers 제목 키워드 기반

ROBOPOL: Social Robotics Meets Vehicular Communications for Cooperative Automated Driving

2025-12-30 · John Pravin Arockiasamy, Andy Comeca, Victoria Yang, Manuel Bied 외 arxiv

On the way toward full autonomy, sharing roads between automated and autonomous vehicles in so-called mixed traffic is unavoidable. Moreover, even if all vehicles on the road were autonomous, pedestrians would still cros…

Autonomous Vehicles

V2X-QA: A Comprehensive Reasoning Dataset and Benchmark for Multimodal Large Language Models in Autonomous Driving Across Ego, Infrastructure, and Cooperative Views

2026-04-03 · Junwei You, Pei Li, Zhuoyu Jiang, Weizhe Tang 외 arxiv

Multimodal large language models (MLLMs) have shown strong potential for autonomous driving, yet existing benchmarks remain largely ego-centric and therefore cannot systematically assess model performance in infrastructu…

Autonomous DrivingQuestion Answering

V2V-LLM: Vehicle-to-Vehicle Cooperative Autonomous Driving with Multi-Modal Large Language Models

2025-02-14 · Hsu-kuang Chiu, Ryo Hachiuma, Chien-Yi Wang, Stephen F. Smith 외

Current autonomous driving vehicles rely mainly on their individual sensors to understand surrounding scenes and plan for future trajectories, which can be unreliable when the sensors are malfunctioning or occluded. To a…

Autonomous DrivingAutonomous VehiclesLarge Language ModelQuestion Answering

Research Challenges and Progress in the End-to-End V2X Cooperative Autonomous Driving Competition

2025-07-29 · Ruiyang Hao, Haibao Yu, Jiaru Zhong, Chuanye Wang 외 arxiv

With the rapid advancement of autonomous driving technology, vehicle-to-everything (V2X) communication has emerged as a key enabler for extending perception range and enhancing driving safety by providing visibility beyo…

Autonomous Driving

V2V-GoT: Vehicle-to-Vehicle Cooperative Autonomous Driving with Multimodal Large Language Models and Graph-of-Thoughts

2025-09-22 · Hsu-kuang Chiu, Ryo Hachiuma, Chien-Yi Wang, Yu-Chiang Frank Wang 외 arxiv

Current state-of-the-art autonomous vehicles could face safety-critical situations when their local sensors are occluded by large nearby objects on the road. Vehicle-to-vehicle (V2V) cooperative autonomous driving has be…

Autonomous VehiclesAutonomous Driving