paper-with-me

Papers

Scenario Generation for Risk-Aware Reinforcement Learning with Probably Approximately Safe Guarantees

2026-06-03 · Mohit Prashant, Arvind Easwaran arxiv

Guaranteeing safety is critical to the deployment of reinforcement learning (RL) agents in the real-world, especially as policies learned using deep RL may demonstrate susceptibility to transition perturbations that result in unknown or unsafe behaviour. A method of policy verification is to construct probabilistic barrier-certificates by sampling policy trajectories with respect to safety constraints, thereby demarcating known safe behaviour from unknown behaviour. Obtaining tight upper and lower bounds on the probability of violation of these constraints may be difficult if the policy is susceptible to transition uncertainty or perturbation that places the agent in insufficiently explored states. To address this, we approximate the distribution of the encountered state-space using a variational autoencoder (VAE) and construct upper and lower-bound barrier-certificates using latent characteristics of states to optimize for regions of known, safe behaviour with high confidence. We frame this in our work as a dual optimization problem where the lower-bound barrier-certificate presents a more conservative estimate of the safe region than the upper-bound barrier-certificate. Sampling states that lie within the set difference of the two during training, i.e. the non-robust region, allows us to tighten the upper and lower bounds to provide sharper probabilistic guarantees on safety. Within our study, we describe the guarantees placed and demonstrate the tightness of our bounds experimentally.

📄 PDF Abstract BibTeX arXiv:2606.04812

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

RA-PbRL: Provably Efficient Risk-Aware Preference-Based Reinforcement Learning

2024-10-31 · Yujie Zhao, Jose Efraim Aguilar Escamill, Weyl Lu, Huazheng Wang

Reinforcement Learning from Human Feedback (RLHF) has recently surged in popularity, particularly for aligning large language models and other AI systems with human intentions. At its core, RLHF can be viewed as a specia…

Autonomous Drivingreinforcement-learningReinforcement Learning

Risk-Controllable Multi-View Diffusion for Driving Scenario Generation

2026-03-12 · Hongyi Lin, Wenxiu Shi, Heye Huang, Dingyi Zhuang 외 arxiv

Generating safety-critical driving scenarios is crucial for evaluating and improving autonomous driving systems, but long-tail risky situations are rarely observed in real-world data and difficult to specify through manu…

Autonomous Driving

Extreme Risk Mitigation in Reinforcement Learning using Extreme Value Theory

2023-08-24 · Karthik Somayaji NS, Yu Wang, Malachi Schram, Jan Drgona 외

Risk-sensitive reinforcement learning (RL) has garnered significant attention in recent years due to the growing interest in deploying RL agents in real-world scenarios. A critical aspect of risk awareness involves model…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Learning A Unified Risk Map for Autonomous Driving in Partially Observable Environments

2026-05-21 · Jie Jia, Yaofeng Su, Zeyu Bao, Yun Hong 외 arxiv

Occlusion-aware prediction remains a critical challenge in autonomous driving due to the inherent uncertainty of unobserved regions. Existing approaches either overestimate risk based on reachable states or struggle to p…

Autonomous Driving

Diffusion-Modeled Reinforcement Learning for Carbon and Risk-Aware Microgrid Optimization

2025-07-22 · Yunyi Zhao, Wei Zhang, Cheng Xiang, Hongyang Du 외 arxiv

This paper introduces DiffCarl, a diffusion-modeled carbon- and risk-aware reinforcement learning algorithm for intelligent operation of multi-microgrid systems. With the growing integration of renewables and increasing …

Reinforcement Learning