paper-with-me

홈 › Papers

Safe and Efficient CAV Lane Changing using Decentralised Safety Shields

2025-04-30 · Bharathkumar Hegde, Melanie Bouroche

Lane changing is a complex decision-making problem for Connected and Autonomous Vehicles (CAVs) as it requires balancing traffic efficiency with safety. Although traffic efficiency can be improved by using vehicular communication for training lane change controllers using Multi-Agent Reinforcement Learning (MARL), ensuring safety is difficult. To address this issue, we propose a decentralised Hybrid Safety Shield (HSS) that combines optimisation and a rule-based approach to guarantee safety. Our method applies control barrier functions to constrain longitudinal and lateral control inputs of a CAV to ensure safe manoeuvres. Additionally, we present an architecture to integrate HSS with MARL, called MARL-HSS, to improve traffic efficiency while ensuring safety. We evaluate MARL-HSS using a gym-like environment that simulates an on-ramp merging scenario with two levels of traffic densities, such as light and moderate densities. The results show that HSS provides a safety guarantee by strictly enforcing a dynamic safety constraint defined on a time headway, even in moderate traffic density that offers challenging lane change scenarios. Moreover, the proposed method learns stable policies compared to the baseline, a state-of-the-art MARL lane change controller without a safety shield. Further policy evaluation shows that our method achieves a balance between safety and traffic efficiency with zero crashes and comparable average speeds in light and moderate traffic densities.

📄 PDF Abstract BibTeX arXiv:2505.01453

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous VehiclesMulti-agent Reinforcement Learning

Similar Papers 제목 키워드 기반

Efficient Dynamic Shielding for Parametric Safety Specifications

2025-05-28 · Davide Corsi, Kaushik Mallik, Andoni Rodriguez, Cesar Sanchez

Shielding has emerged as a promising approach for ensuring safety of AI-controlled autonomous systems. The algorithmic goal is to compute a shield, which is a runtime safety enforcement tool that needs to monitor and int…

Robot Navigation

Contract-Based Compositional Shielding for Safe Multi-Agent Reinforcement Learning

2026-06-12 · Omar Adalat, Edwin Hamel-De le Court, Francesco Belardinelli arxiv

Safe coordination problems surface in multi-agent reinforcement learning when global safety cannot be enforced by any agent unilaterally: the admissibility of one agent's action may depend on the dynamics of other agents…

Multi-agent Reinforcement Learning

Safety Shielding under Delayed Observation

2023-07-05 · Filip Cano Córdoba, Alexander Palmisano, Martin Fränzle, Roderick Bloem 외

Agents operating in physical environments need to be able to handle delays in the input and output signals since neither data transmission nor sensing or actuating the environment are instantaneous. Shields are correct-b…

Autonomous Driving

Shields to Guarantee Probabilistic Safety in MDPs

2026-05-11 · Linus Heck, Filip Macák, Roman Andriushchenko, Milan Češka 외 arxiv

Shielding is a prominent model-based technique to ensure safety of autonomous agents. Classical shielding aims to ensure that nothing bad ever happens and comes with strong guarantees about safety and maximal permissiven…

It's Time to Play Safe: Shield Synthesis for Timed Systems

2020-06-30 · Roderick Bloem, Peter Gjøl Jensen, Bettina Könighofer, Kim Guldstrand Larsen 외

Erroneous behaviour in safety critical real-time systems may inflict serious consequences. In this paper, we show how to synthesize timed shields from timed safety properties given as timed automata. A timed shield enfor…