paper-with-me

Papers

Learning from Mistakes: Post-Training for Driving VLA with Takeover Data

2026-03-16 · Yinfeng Gao, Deqing Liu, Qichao Zhang, Yupeng Zheng, Haochen Tian, Guang Li, Hangjun Ye, Long Chen, Da-Wei Ding, Dongbin Zhao arxiv

Current Vision-Language-Action (VLA) paradigms in end-to-end autonomous driving rely on offline training from static datasets, leaving them vulnerable to distribution shift. Recent post-training methods use takeover data to mitigate this by augmenting the dataset with high-quality expert takeover samples, yet they suffer from two key limitations: supervision restricted to the period after the takeover moments leads to policies with limited safety margins, and passive preference optimization lacks active exploration for optimal performance. In this paper, we propose TakeVLA, a novel VLA post-training framework that overcomes these shortcomings through two complementary innovations. First, we introduce pre-takeover language supervision, which allows the VLA to learn from mistakes proactively. By explicitly teaching the model about what to do in error-prone situations, we cultivate a precautionary mindset that anticipates hazards early and substantially enlarges safety margins. Second, we propose Scenario Dreaming, a reinforcement fine-tuning paradigm that operates in reconstruceted takeover scenarios, encouraging active exploration beyond mere preference fitting. Experiments on the Bench2Drive benchmark demonstrate that TakeVLA achieves state-of-the-art closed-loop performance, surpassing the strong VLA baseline SimLingo by 4.93 in driving score, with an enhanced safety margin as evidenced by an 11.76% increase in average TTC.

📄 PDF Abstract BibTeX arXiv:2603.14972

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous Driving

Results from the Paper

RankTaskDatasetModelMetrics
#92 Bench2Drive Bench2Drive TakeVLA Driving Score: 4.93

Similar Papers 제목 키워드 기반

TakeAD: Preference-based Post-optimization for End-to-end Autonomous Driving with Expert Takeover Data

2025-12-19 · Deqing Liu, Yinfeng Gao, Deheng Qian, Qichao Zhang 외 arxiv

Existing end-to-end autonomous driving methods typically rely on imitation learning (IL) but face a key challenge: the misalignment between open-loop training and closed-loop deployment. This misalignment often triggers …

Autonomous Driving

Examining the Effects of Emotional Valence and Arousal on Takeover Performance in Conditionally Automated Driving

2020-01-13 · Na Du, Feng Zhou, Elizabeth Pulver, Dawn M. Tilbury 외

In conditionally automated driving, drivers have difficulty in takeover transitions as they become increasingly decoupled from the operational level of driving. Factors influencing takeover performance, such as takeover …

A quantitative model of takeover request time budget for conditionally automated driving

2024-08-28 · Foghor Tanshi, Dirk Söffker

In conditional automation, the automated driving system assumes full control and only issues a takeover request to a human driver to resume driving in critical situations. Previous studies have concluded that the time bu…

Predicting Driver Takeover Time in Conditionally Automated Driving

2021-07-20 · Jackie Ayoub, Na Du, X. Jessie Yang, Feng Zhou

It is extremely important to ensure a safe takeover transition in conditionally automated driving. One of the critical factors that quantifies the safe takeover transition is takeover time. Previous studies identified th…

DeepTake: Prediction of Driver Takeover Behavior using Multimodal Data

2020-12-31 · Erfan Pakdamanian, Shili Sheng, Sonia Baee, Seongkook Heo 외

Automated vehicles promise a future where drivers can engage in non-driving tasks without hands on the steering wheels for a prolonged period. Nevertheless, automated vehicles may still need to occasionally hand the cont…