paper-with-me

Papers

Robust RL with LLM-Driven Data Synthesis and Policy Adaptation for Autonomous Driving

2024-10-16 · Sihao Wu, Jiaxu Liu, Xiangyu Yin, Guangliang Cheng, Xingyu Zhao, Meng Fang, Xinping Yi, Xiaowei Huang

The integration of Large Language Models (LLMs) into autonomous driving systems demonstrates strong common sense and reasoning abilities, effectively addressing the pitfalls of purely data-driven methods. Current LLM-based agents require lengthy inference times and face challenges in interacting with real-time autonomous driving environments. A key open question is whether we can effectively leverage the knowledge from LLMs to train an efficient and robust Reinforcement Learning (RL) agent. This paper introduces RAPID, a novel \underline{\textbf{R}}obust \underline{\textbf{A}}daptive \underline{\textbf{P}}olicy \underline{\textbf{I}}nfusion and \underline{\textbf{D}}istillation framework, which trains specialized mix-of-policy RL agents using data synthesized by an LLM-based driving agent and online adaptation. RAPID features three key designs: 1) utilization of offline data collected from an LLM agent to distil expert knowledge into RL policies for faster real-time inference; 2) introduction of robust distillation in RL to inherit both performance and robustness from LLM-based teacher; and 3) employment of a mix-of-policy approach for joint decision decoding with a policy adapter. Through fine-tuning via online environment interaction, RAPID reduces the forgetting of LLM knowledge while maintaining adaptability to different tasks. Extensive experiments demonstrate RAPID's capability to effectively integrate LLM knowledge into scaled-down RL policies in an efficient, adaptable, and robust way. Code and checkpoints will be made publicly available upon acceptance.

📄 PDF Abstract BibTeX arXiv:2410.12568

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingCommon Sense ReasoningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Optimal by Design: Model-Driven Synthesis of Adaptation Strategies for Autonomous Systems

2020-01-16 · Yehia Elrakaiby, Paola Spoletini, Bashar Nuseibeh

Many software systems have become too large and complex to be managed efficiently by human administrators, particularly when they operate in uncertain and dynamic environments and require frequent changes. Requirements-d…

LOGIGEN: Logic-Driven Generation of Verifiable Agentic Tasks

2026-02-28 · Yucheng Zeng, Weipeng Lu, Linyun Liu, Shupeng Li 외 arxiv

The evolution of Large Language Models (LLMs) from static instruction-followers to autonomous agents necessitates operating within complex, stateful environments to achieve precise state-transition objectives. However, t…

Reinforcement Learning

GreenIQ: A Deep Search Platform for Comprehensive Carbon Market Analysis and Automated Report Generation

2025-03-20 · Bisola Faith Kayode, Akinyemi Sadeeq Akintola, Oluwole Fagbohun, Egonna Anaesiuba-Bristol 외

This study introduces GreenIQ, an AI-powered deep search platform designed to revolutionise carbon market intelligence through autonomous analysis and automated report generation. Carbon markets operate across diverse re…

Information Retrieval

Multi-Objective Reinforcement Learning for Adaptive Personalized Autonomous Driving

2025-05-08 · Hendrik Surmann, Jorge de Heuvel, Maren Bennewitz

Human drivers exhibit individual preferences regarding driving style. Adapting autonomous vehicles to these preferences is essential for user trust and satisfaction. However, existing end-to-end driving approaches often …

Autonomous DrivingAutonomous VehiclesCollision AvoidanceMulti-Objective Reinforcement Learning+2

Cyber Shadows: Neutralizing Security Threats with AI and Targeted Policy Measures

2025-01-03 · Marc Schmitt, Pantelis Koutroumpis

The digital age, driven by the AI revolution, brings significant opportunities but also conceals security threats, which we refer to as cyber shadows. These threats pose risks at individual, organizational, and societal …

Intrusion Detection