paper-with-me

홈 › Papers

Found-RL: foundation model-enhanced reinforcement learning for autonomous driving

2026-02-11 · Yansong Qu, Zihao Sheng, Zilin Huang, Jiancong Chen, Yuhao Luo, Tianyi Wang, Yiheng Feng, Samuel Labi, Sikai Chen arxiv

Reinforcement Learning (RL) has emerged as a dominant paradigm for end-to-end autonomous driving (AD). However, RL suffers from sample inefficiency and a lack of semantic interpretability in complex scenarios. Foundation Models, particularly Vision-Language Models (VLMs), can mitigate this by offering rich, context-aware knowledge, yet their high inference latency hinders deployment in high-frequency RL training loops. To bridge this gap, we present Found-RL, a platform tailored to efficiently enhance RL for AD using foundation models. A core innovation is the asynchronous batch inference framework, which decouples heavy VLM reasoning from the simulation loop, effectively resolving latency bottlenecks to support real-time learning. We introduce diverse supervision mechanisms: Value-Margin Regularization (VMR) and Advantage-Weighted Action Guidance (AWAG) to effectively distill expert-like VLM action suggestions into the RL policy. Additionally, we adopt high-throughput CLIP for dense reward shaping. We address CLIP's dynamic blindness via Conditional Contrastive Action Alignment, which conditions prompts on discretized speed/command and yields a normalized, margin-based bonus from context-specific action-anchor scoring. Found-RL provides an end-to-end pipeline for fine-tuned VLM integration and shows that a lightweight RL model can achieve near-VLM performance compared with billion-parameter VLMs while sustaining real-time inference (approx. 500 FPS). Code, data, and models will be publicly available at https://github.com/ys-qu/found-rl.

📄 PDF Abstract BibTeX arXiv:2602.10458

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningAutonomous Driving

Similar Papers 제목 키워드 기반

Panacea+: Panoramic and Controllable Video Generation for Autonomous Driving

2024-08-14 · Yuqing Wen, Yucheng Zhao, Yingfei Liu, Binyuan Huang 외

The field of autonomous driving increasingly demands high-quality annotated video training data. In this paper, we propose Panacea+, a powerful and universally applicable framework for generating video data in driving sc…

3D Object Detection3D Object TrackingAutonomous DrivingLane Detection+6

Sky-Drive: A Distributed Multi-Agent Simulation Platform for Human-AI Collaborative and Socially-Aware Future Transportation

2025-04-25 · Zilin Huang, Zihao Sheng, Zhengyang Wan, Yansong Qu 외

Recent advances in autonomous system simulation platforms have significantly enhanced the safe and scalable testing of driving policies. However, existing simulators do not yet fully meet the needs of future transportati…

Applications of Large Scale Foundation Models for Autonomous Driving

2023-11-20 · Yu Huang, Yue Chen, Zhu Li

Since DARPA Grand Challenges (rural) in 2004/05 and Urban Challenges in 2007, autonomous driving has been the most active field of AI applications. Recently powered by large language models (LLMs), chat systems, such as …

Autonomous Driving

VLP: Vision Language Planning for Autonomous Driving

2024-01-10 · CVPR 2024 1 · Chenbin Pan, Burhaneddin Yaman, Tommaso Nesti, Abhirup Mallik 외

Autonomous driving is a complex and challenging task that aims at safe motion planning through scene understanding and reasoning. While vision-only autonomous driving methods have recently achieved notable performance, t…

Autonomous DrivingMotion PlanningScene Understanding

Knowledge Integration Strategies in Autonomous Vehicle Prediction and Planning: A Comprehensive Survey

2025-02-13 · Kumar Manas, Adrian Paschke

This comprehensive survey examines the integration of knowledge-based approaches into autonomous driving systems, with a focus on trajectory prediction and planning. We systematically review methodologies for incorporati…

Autonomous DrivingFormal LogicSurveyTrajectory Prediction