paper-with-me

Papers

Urban Driver: Learning to Drive from Real-world Demonstrations Using Policy Gradients

2021-09-27 · Oliver Scheel, Luca Bergamini, Maciej Wołczyk, Błażej Osiński, Peter Ondruska

In this work we are the first to present an offline policy gradient method for learning imitative policies for complex urban driving from a large corpus of real-world demonstrations. This is achieved by building a differentiable data-driven simulator on top of perception outputs and high-fidelity HD maps of the area. It allows us to synthesize new driving experiences from existing demonstrations using mid-level representations. Using this simulator we then train a policy network in closed-loop employing policy gradients. We train our proposed method on 100 hours of expert demonstrations on urban roads and show that it learns complex driving policies that generalize well and can perform a variety of driving maneuvers. We demonstrate this in simulation as well as deploy our model to self-driving vehicles in the real-world. Our method outperforms previously demonstrated state-of-the-art for urban driving scenarios -- all this without the need for complex state perturbations or collecting additional on-policy data during training. We make code and data publicly available.

📄 PDF Abstract BibTeX arXiv:2109.13333

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Driving Style Alignment for LLM-powered Driver Agent

2024-03-17 · Ruoxuan Yang, Xinyue Zhang, Anais Fernandez-Laaksonen, Xin Ding 외

Recently, LLM-powered driver agents have demonstrated considerable potential in the field of autonomous driving, showcasing human-like reasoning and decision-making abilities.However, current research on aligning driver …

Autonomous DrivingDecision Making

A Kalman Filter-Based Disturbance Observer for Steer-by-Wire Systems

2025-12-29 · Nikolai Beving, Jonas Marxen, Steffen Mueller, Johannes Betz arxiv

Steer-by-Wire systems replace mechanical linkages, which provide benefits like weight reduction, design flexibility, and compatibility with autonomous driving. However, they are susceptible to high-frequency disturbances…

Autonomous Driving

Inverse Resource Rational Based Stochastic Driver Behavior Model

2022-07-14 · Mehmet Ozkan, Yao Ma

Human drivers have limited and time-varying cognitive resources when making decisions in real-world traffic scenarios, which often leads to unique and stochastic behaviors that can not be explained by perfect rationality…

modelModel Predictive Control

GARLIC: GPT-Augmented Reinforcement Learning with Intelligent Control for Vehicle Dispatching

2024-08-19 · Xiao Han, Zijian Zhang, Xiangyu Zhao, Yuanshao Zhu 외

As urban residents demand higher travel quality, vehicle dispatch has become a critical component of online ride-hailing services. However, current vehicle dispatch systems struggle to navigate the complexities of urban …

Navigatereinforcement-learningReinforcement Learning

Inverse Reinforcement Learning Based Stochastic Driver Behavior Learning

2021-07-01 · Mehmet Fatih Ozkan, Abishek Joseph Rocque, Yao Ma

Drivers have unique and rich driving behaviors when operating vehicles in traffic. This paper presents a novel driver behavior learning approach that captures the uniqueness and richness of human driver behavior in reali…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)