paper-with-me

홈 › Papers

Causality-Aware End-to-End Autonomous Driving via Ego-Centric Joint Scene Modeling

2026-05-13 · Seokha Moon, Minseung Lee, Joon Seo, Jinkyu Kim, Jungbeom Lee arxiv

End-to-end autonomous driving, which bypasses traditional modular pipelines by directly predicting future trajectories from sensor inputs, has recently achieved substantial progress. However, existing methods often overlook the causal inter-dependencies in ego-vehicle planning, ignoring the reciprocal relations between the ego vehicle and surrounding agents. This causal oversight leads to inconsistent and unreliable trajectory predictions, especially in interaction-critical scenarios where ego decisions and neighboring agent behaviors must be reasoned about jointly. To address this limitation, we propose CaAD, a Causality-aware end-to-end Autonomous Driving framework that captures these dependencies within a shared latent scene representation. First, we propose an ego-centric joint-causal modeling module that builds on the marginal prediction branch, and learns causal dependencies between the ego vehicle and interaction-relevant agents. Second, we employ a causality-aware policy alignment stage implemented with joint-mode embeddings to align the stochastic ego policy with planning-oriented closed-loop feedback computed from surrounding traffic and map context. On the Bench2Drive and NAVSIM benchmarks, CaAD demonstrates strong closed-loop planning performance, achieving a Driving Score of 87.53 and Success Rate of 71.81 on Bench2Drive, and a PDMS of 91.1 on NAVSIM. The project page is available at https://moonseokha.github.io/CaAD/.

📄 PDF Abstract BibTeX arXiv:2605.13646

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous Driving

Results from the Paper

RankTaskDatasetModelMetrics
#4 Bench2Drive Bench2Drive CaAD Driving Score: 87.53

Similar Papers 제목 키워드 기반

Exploring the Causality of End-to-End Autonomous Driving

2024-07-09 · Jiankun Li, Hao Li, JiangJiang Liu, Zhikang Zou 외

Deep learning-based models are widely deployed in autonomous driving areas, especially the increasingly noticed end-to-end solutions. However, the black-box property of these models raises concerns about their trustworth…

Autonomous Drivingcounterfactual

FeaXDrive: Feasibility-aware Trajectory-Centric Diffusion Planning for End-to-End Autonomous Driving

2026-04-14 · Baoyun Wang, Zhuoren Li, Ran Yu, Yu Che 외 arxiv

End-to-end diffusion planning has shown strong potential for autonomous driving, but the physical feasibility of generated trajectories remains insufficiently addressed. In particular, generated trajectories may exhibit …

Autonomous Driving

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving

2025-12-04 · Yihong Tang, Haicheng Liao, Tong Nie, Junlin He 외 arxiv

End-to-end autonomous driving (AD) systems increasingly adopt vision-language-action (VLA) models, yet they typically ignore the passenger's emotional state, which is central to comfort and AD acceptance. We introduce Op…

Autonomous DrivingSpatial ReasoningVisual Grounding

Where Does the Answer Come From? Benchmarking View-Level Visual Evidence Identification in Multi-View MLLMs for Autonomous Driving

2026-06-08 · Yimu Wang, Yee Man Choi, Barry Zhang, Mozhgan Nasr Azadani 외 arxiv

Multimodal large language models (MLLMs) achieve strong results on visual reasoning benchmarks, but answer accuracy alone does not indicate whether a model relied on the correct visual evidence. This gap is particularly …

Visual Question AnsweringAutonomous DrivingVisual Reasoning

Learning to Drive is a Free Gift: Large-Scale Label-Free Autonomy Pretraining from Unposed In-The-Wild Videos

2026-02-25 · Matthew Strong, Wei-Jer Chang, Quentin Herau, Jiezhi Yang 외 arxiv

Ego-centric driving videos available online provide an abundant source of visual data for autonomous driving, yet their lack of annotations makes it difficult to learn representations that capture both semantic structure…

Semantic SegmentationAutonomous Driving