paper-with-me

홈 › Papers

VLP: Vision Language Planning for Autonomous Driving

2024-01-10 · CVPR 2024 1 · Chenbin Pan, Burhaneddin Yaman, Tommaso Nesti, Abhirup Mallik, Alessandro G Allievi, Senem Velipasalar, Liu Ren

Autonomous driving is a complex and challenging task that aims at safe motion planning through scene understanding and reasoning. While vision-only autonomous driving methods have recently achieved notable performance, through enhanced scene understanding, several key issues, including lack of reasoning, low generalization performance and long-tail scenarios, still need to be addressed. In this paper, we present VLP, a novel Vision-Language-Planning framework that exploits language models to bridge the gap between linguistic understanding and autonomous driving. VLP enhances autonomous driving systems by strengthening both the source memory foundation and the self-driving car's contextual understanding. VLP achieves state-of-the-art end-to-end planning performance on the challenging NuScenes dataset by achieving 35.9\% and 60.5\% reduction in terms of average L2 error and collision rates, respectively, compared to the previous best method. Moreover, VLP shows improved performance in challenging long-tail scenarios and strong generalization capabilities when faced with new urban environments.

📄 PDF Abstract BibTeX arXiv:2401.05577

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingMotion PlanningScene Understanding

Similar Papers 제목 키워드 기반

LLaViDA: A Large Language Vision Driving Assistant for Explicit Reasoning and Enhanced Trajectory Planning

2025-12-20 · Yudong Liu, Spencer Hallyburton, Jiwoo Kim, Yueqian Lin 외 arxiv

Trajectory planning is a fundamental yet challenging component of autonomous driving. End-to-end planners frequently falter under adverse weather, unpredictable human behavior, or complex road layouts, primarily because …

Scene UnderstandingTrajectory PlanningAutonomous Driving

AppleVLM: End-to-end Autonomous Driving with Advanced Perception and Planning-Enhanced Vision-Language Models

2026-02-04 · Yuxuan Han, Kunyuan Wu, Qianyi Shao, Renxiang Xiao 외 arxiv

End-to-end autonomous driving has emerged as a promising paradigm integrating perception, decision-making, and control within a unified learning framework. Recently, Vision-Language Models (VLMs) have gained significant …

Autonomous Driving

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model

2024-12-13 · Songyan Zhang, Wenhui Huang, Zihui Gao, Hao Chen 외

The emergence of general human knowledge and impressive logical reasoning capacity in rapidly progressed vision-language models (VLMs) have driven increasing interest in applying VLMs to high-level autonomous driving tas…

Autonomous DrivingDecision MakingLanguage ModelingLanguage Modelling+4

Generative Planning with 3D-vision Language Pre-training for End-to-End Autonomous Driving

2025-01-15 · Tengpeng Li, Hanli Wang, Xianfei Li, Wenlong Liao 외

Autonomous driving is a challenging task that requires perceiving and understanding the surrounding environment for safe trajectory planning. While existing vision-based end-to-end models have achieved promising results,…

Autonomous DrivingTrajectory Planning

Distilling Multi-modal Large Language Models for Autonomous Driving

2025-01-16 · CVPR 2025 1 · Deepti Hegde, Rajeev Yasarla, Hong Cai, Shizhong Han 외

Autonomous driving demands safe motion planning, especially in critical "long-tail" scenarios. Recent end-to-end autonomous driving systems leverage large language models (LLMs) as planners to improve generalizability to…

Autonomous DrivingMotion PlanningWorld Knowledge