paper-with-me

Papers

VLADriveBench: Evaluating CoT-Action Relationship in VLA for Autonomous Driving

2026-06-10 · Thach Nguyen, Danhua Guo, Tom Lampo, Fei Wu, Burhan Yaman arxiv

Vision-language-action (VLA) models generate chain-of-thought (CoT) reasoning alongside driving trajectories, but existing benchmarks evaluate only trajectory quality and do not assess whether the CoT is relevant, consistent, or causally connected to the driving action. We introduce VLADriveBench, a framework that combines observational metrics (mentioning, hallucination, contradiction, action alignment) with a CoT intervention protocol to provide complementary views of the CoT-action relationship. Applying VLADriveBench to three models across two architectures, we find that the two analyses can diverge sharply: ORION scores highest on observational alignment yet its CoT is epiphenomenal, while Alpamayo v1.5 scores lower yet its CoT is strongly causal, with visual salience gating the extent of CoT influence.

📄 PDF Abstract BibTeX arXiv:2606.12706

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous Driving

Similar Papers 제목 키워드 기반

Driving in Corner Case: A Real-World Adversarial Closed-Loop Evaluation Platform for End-to-End Autonomous Driving

2025-12-18 · Jiaheng Geng, Jiatong Du, Xinyu Zhang, Ye Li 외 arxiv

Safety-critical corner cases, difficult to collect in the real world, are crucial for evaluating end-to-end autonomous driving. Adversarial interaction is an effective method to generate such safety-critical corner cases…

Autonomous Driving

ReasonBreak: Probing Vulnerabilities in Reasoning-Enabled Vision-Language-Action Models for Autonomous Driving

2026-05-27 · Mohammadreza Teymoorianfard, Jean-Philippe Monteuuis, Jonathan Petit, Amir Houmansadr arxiv

Vision-Language-Action (VLA) models with integrated reasoning have been proposed for end-to-end autonomous driving, assuming a tight coupling between reasoning and trajectory generation. However, the robustness of such s…

Autonomous Driving

Large Language Models for Autonomous Driving (LLM4AD): Concept, Benchmark, Experiments, and Challenges

2024-10-20 · Can Cui, Yunsheng Ma, Zichong Yang, Yupeng Zhou 외

With the broader usage and highly successful development of Large Language Models (LLMs), there has been a growth of interest and demand for applying LLMs to autonomous driving technology. Driven by their natural languag…

Autonomous DrivingDecision MakingInstruction FollowingNatural Language Understanding+2

An interactive enhanced driving dataset for autonomous driving

2026-02-24 · Haojie Feng, Peizhi Zhang, Mengjie Tian, Xinrui Zhang 외 arxiv

The evolution of autonomous driving towards full automation demands robust interactive capabilities; however, the development of Vision-Language-Action (VLA) models is constrained by the sparsity of interactive scenarios…

Autonomous Driving

VaViM and VaVAM: Autonomous Driving through Video Generative Modeling

2025-02-21 · Florent Bartoccioni, Elias Ramzi, Victor Besnier, Shashanka Venkataramanan 외

We explore the potential of large-scale generative video models for autonomous driving, introducing an open-source auto-regressive video model (VaViM) and its companion video-action model (VaVAM) to investigate how video…

Autonomous DrivingImitation Learning