paper-with-me

Papers

AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding

2026-07-09 · Siddharth Damodharan, Radhika Gupta, Ali Alshami, Ryan Rabinowitz, Jugal Kalita arxiv

Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as scene understanding, decision making, trajectory prediction, and visual question answering. However, evaluating whether these models can reliably reason about safety-critical incidents remains challenging. To address this gap, we present AUTOPILOT-VQA, an incident-centric visual question answering benchmark for dashcam video understanding. The dataset evaluates different systems through structured questions designed around real-world driving incidents and near-incidents. The benchmark covers diverse safety-relevant categories, including weather and lighting conditions, traffic environment, road layout, road surface state, signage, involved entities, accident occurrence, impact location, and avoidability-related reasoning. By requiring models to answer grounded questions about both contextual scene properties and event-level incident details, AUTOPILOT-VQA moves beyond object recognition toward temporally grounded, safety-aware reasoning. The dataset is released as part of the AUTOPILOT CVPR 2026 competition and provides a standardized benchmark for assessing the reliability of autonomous driving systems in different scenarios. Our benchmark support developments for more interpretable, robust, and safety-conscious vision-language systems for real-world autonomous driving.

📄 PDF Abstract BibTeX arXiv:2607.08745

Code (0)

등록된 구현이 없습니다.

Tasks

Visual Question AnsweringTrajectory PredictionScene UnderstandingAutonomous Driving

Similar Papers 제목 키워드 기반

Tesla's Autopilot: Ethics and Tragedy

2024-09-25 · Aravinda Jatavallabha

This case study delves into the ethical ramifications of an incident involving Tesla's Autopilot, emphasizing Tesla Motors' moral responsibility. Using a seven-step ethical decision-making process, it examines user behav…

Decision MakingEthics

Cognitive Kernel: An Open-source Agent System towards Generalist Autopilots

2024-09-16 · Hongming Zhang, Xiaoman Pan, Hongwei Wang, Kaixin Ma 외

We introduce Cognitive Kernel, an open-source agent system towards the goal of generalist autopilots. Unlike copilot systems, which primarily rely on users to provide essential state information (e.g., task descriptions)…

Management

A data-centric weak supervised learning for highway traffic incident detection

2021-12-17 · Yixuan Sun, Tanwi Mallick, Prasanna Balaprakash, Jane Macfarlane

Using the data from loop detector sensors for near-real-time detection of traffic incidents in highways is crucial to averting major traffic congestion. While recent supervised machine learning methods offer solutions to…

Uncertainty Quantification

Autopilot-Preserving Residual Q-Learning with HJB-Inspired Finite-Action Risk Filtering for Fixed-Wing UAV Command Supervision

2026-05-31 · Mehmet Iscan, Batuhan Temiz arxiv

A fixed-wing UAV must hold airspeed, altitude, and heading references under wind, gusts, and turbulence, channels coupled so that correcting one can degrade another. Classical autopilots stabilize the airframe well but a…

Experimental Flight Testing of a Fault-Tolerant Adaptive Autopilot for Fixed-Wing Aircraft

2022-10-24 · Joonghyun Lee, John Spencer, Siyuan Shao, Juan Augusto Paredes 외

This paper presents an adaptive autopilot for fixed-wing aircraft and compares its performance with a fixed-gain autopilot. The adaptive autopilot is constructed by augmenting the autopilot architecture with adaptive con…