paper-with-me

홈 › Papers

Ensuring Safety in an Uncertain Environment: Constrained MDPs via Stochastic Thresholds

2025-04-07 · Qian Zuo, Fengxiang He

This paper studies constrained Markov decision processes (CMDPs) with constraints against stochastic thresholds, aiming at the safety of reinforcement learning in unknown and uncertain environments. We leverage a Growing-Window estimator sampling from interactions with the uncertain and dynamic environment to estimate the thresholds, based on which we design Stochastic Pessimistic-Optimistic Thresholding (SPOT), a novel model-based primal-dual algorithm for multiple constraints against stochastic thresholds. SPOT enables reinforcement learning under both pessimistic and optimistic threshold settings. We prove that our algorithm achieves sublinear regret and constraint violation; i.e., a reward regret of $\tilde{\mathcal{O}}(\sqrt{T})$ while allowing an $\tilde{\mathcal{O}}(\sqrt{T})$ constraint violation over $T$ episodes. The theoretical guarantees show that our algorithm achieves performance comparable to that of an approach relying on fixed and clear thresholds. To the best of our knowledge, SPOT is the first reinforcement learning algorithm that realises theoretical guaranteed performance in an uncertain environment where even thresholds are unknown.

📄 PDF Abstract BibTeX arXiv:2504.04973

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

Lyapunov Robust Constrained-MDPs: Soft-Constrained Robustly Stable Policy Optimization under Model Uncertainty

2021-08-05 · Reazul Hasan Russel, Mouhacine Benosman, Jeroen van Baar, Radu Corcodel

Safety and robustness are two desired properties for any reinforcement learning algorithm. CMDPs can handle additional safety constraints and RMDPs can perform well under model uncertainties. In this paper, we propose to…

reinforcement-learningReinforcement Learning (RL)

Enhanced Safety in Autonomous Driving: Integrating Latent State Diffusion Model for End-to-End Navigation

2024-07-08 · Detian Chu, Linyuan Bai, Jianuo Huang, Zhenlong Fang 외

With the advancement of autonomous driving, ensuring safety during motion planning and navigation is becoming more and more important. However, most end-to-end planning methods suffer from a lack of safety. This research…

Autonomous DrivingDecision MakingMotion PlanningSafe Exploration

Flipping-based Policy for Chance-Constrained Markov Decision Processes

2024-10-09 · Xun Shen, Shuo Jiang, Akifumi Wachi, Kaumune Hashimoto 외

Safe reinforcement learning (RL) is a promising approach for many real-world decision-making problems where ensuring safety is a critical necessity. In safe RL research, while expected cumulative safety constraints (ECSC…

Reinforcement Learning (RL)Safe Reinforcement Learning

Provably Efficient Sample Complexity for Robust CMDP

2025-11-10 · Sourav Ganguly, Arnob Ghosh arxiv

We study the problem of learning policies that maximize cumulative reward while satisfying safety constraints, even when the real environment differs from a simulator or nominal model. We focus on robust constrained Mark…

Lyapunov-based uncertainty-aware safe reinforcement learning

2021-07-29 · Ashkan B. Jeddi, Nariman L. Dehghani, Abdollah Shafieezadeh

Reinforcement learning (RL) has shown a promising performance in learning optimal policies for a variety of sequential decision-making tasks. However, in many real-world RL problems, besides optimizing the main objective…

Autonomous DrivingDecision Makingreinforcement-learningReinforcement Learning+4