paper-with-me

Papers

OpenAD: Open-World Autonomous Driving Benchmark for 3D Object Detection

2024-11-26 · Zhongyu Xia, Jishuo Li, Zhiwei Lin, Xinhao Wang, Yongtao Wang, Ming-Hsuan Yang

Open-world autonomous driving encompasses domain generalization and open-vocabulary. Domain generalization refers to the capabilities of autonomous driving systems across different scenarios and sensor parameter configurations. Open vocabulary pertains to the ability to recognize various semantic categories not encountered during training. In this paper, we introduce OpenAD, the first real-world open-world autonomous driving benchmark for 3D object detection. OpenAD is built on a corner case discovery and annotation pipeline integrating with a multimodal large language model (MLLM). The proposed pipeline annotates corner case objects in a unified format for five autonomous driving perception datasets with 2000 scenarios. In addition, we devise evaluation methodologies and evaluate various 2D and 3D open-world and specialized models. Moreover, we propose a vision-centric 3D open-world object detection baseline and further introduce an ensemble method by fusing general and specialized models to address the issue of lower precision in existing open-world methods for the OpenAD benchmark. Annotations, toolkit code, and all evaluation codes will be released.

📄 PDF Abstract BibTeX arXiv:2411.17761

Code (1)

VDIGPKU/OpenAD 공식 구현 pytorch

Tasks

3D Object DetectionAutonomous DrivingDomain GeneralizationLanguage ModelingLanguage ModellingLarge Language ModelMultimodal Large Language Modelobject-detectionObject DetectionOpen World Object Detection

Similar Papers 제목 키워드 기반

Open-Vocabulary Affordance Detection in 3D Point Clouds

2023-03-04 · Toan Nguyen, Minh Nhat Vu, An Vuong, Dzung Nguyen 외

Affordance detection is a challenging problem with a wide variety of robotic applications. Traditional affordance detection methods are limited to a predefined set of affordance labels, hence potentially restricting the …

Affordance Detection

Multi-Agent Embodied Autonomous Driving (MAEAD): From V2X Information Exchange to Shared World Models

2026-06-11 · Senkang Hu, Zhengru Fang, Yihang Tao, Zihan Fang 외 arxiv

Autonomous driving is shifting from isolated vehicle intelligence toward multi-agent embodied systems that share perception, infer intent, and coordinate action under uncertainty. This survey examines this transition thr…

Autonomous Driving

Open-World Panoptic Segmentation

2024-12-17 · Matteo Sodano, Federico Magistri, Jens Behley, Cyrill Stachniss

Perception is a key building block of autonomously acting vision systems such as autonomous vehicles. It is crucial that these systems are able to understand their surroundings in order to operate safely and robustly. Ad…

Autonomous DrivingAutonomous VehiclesPanoptic SegmentationSegmentation+1

X-Driver: Explainable Autonomous Driving with Vision-Language Models

2025-05-08 · Wei Liu, Jiyuan Zhang, Binxiong Zheng, Yufeng Hu 외

End-to-end autonomous driving has advanced significantly, offering benefits such as system simplicity and stronger driving performance in both open-loop and closed-loop settings than conventional pipelines. However, exis…

Autonomous DrivingBench2DriveDecision Making

One-Stage Object Detectors in Autonomous Driving

2026-08-19 · Jonel Roman, Ryan Sirjue, Peter Nguyen, Daniel Krutky 외 arxiv

Autonomous vehicles depend on fast and reliable perception systems to detect surrounding vehicles, pedestrians, cyclists, traffic signs, and other road objects in real time. This paper presents a comprehensive survey and…

Autonomous VehiclesAutonomous Driving