Learning to Drive Safely with Hybrid Options
Out of the many deep reinforcement learning approaches for autonomous driving, only few make use of the options (or skills) framework. That is surprising, as this framework is naturally suited for hierarchical control applications in general, and autonomous driving tasks in specific. Therefore, in this work the options framework is applied and tailored to autonomous driving tasks on highways. More specifically, we define dedicated options for longitudinal and lateral manoeuvres with embedded safety and comfort constraints. This way, prior domain knowledge can be incorporated into the learning process and the learned driving behaviour can be constrained more easily. We propose several setups for hierarchical control with options and derive practical algorithms following state-of-the-art reinforcement learning techniques. By separately selecting actions for longitudinal and lateral control, the introduced policies over combined and hybrid options obtain the same expressiveness and flexibility that human drivers have, while being easier to interpret than classical policies over continuous actions. Of all the investigated approaches, these flexible policies over hybrid options perform the best under varying traffic conditions, outperforming the baseline policies over actions.
Code (0)
등록된 구현이 없습니다.
Tasks
Reinforcement LearningAutonomous DrivingSimilar Papers 제목 키워드 기반
Pricing vulnerable options in a hybrid credit risk model driven by Heston-Nandi GARCH processes
This paper proposes a hybrid credit risk model, in closed form, to price vulnerable options with stochastic volatility. The distinctive features of the model are threefold. First, both the underlying and the option issue…
Identifying Distinct, Effective Treatments for Acute Hypotension with SODA-RL: Safely Optimized Diverse Accurate Reinforcement Learning
Hypotension in critical care settings is a life-threatening emergency that must be recognized and treated early. While fluid bolus therapy and vasopressors are common treatments, it is often unclear which interventions t…
Reinforcement LearningToward an efficient hybrid method for pricing barrier options on assets with stochastic volatility
We combine the one-dimensional Monte Carlo simulation and the semi-analytical one-dimensional heat potential method to design an efficient technique for pricing barrier options on assets with correlated stochastic volati…
To Compare, or Not to Compare: On Methodological Practices in Evaluating Social Bias
As Large Language Models are increasingly deployed in critical applications, robustly evaluating their social biases is paramount. However, the current literature suffers from widespread methodological fragmentation, whi…
Net Buying Pressure and the Information in Bitcoin Option Trades
How do supply and demand from informed traders drive market prices of bitcoin options? Deribit options tick-level data supports the limits-to-arbitrage hypothesis about the market maker's supply. The main demand-side eff…