paper-with-me

Papers

Sampling-based Safe Reinforcement Learning for Nonlinear Dynamical Systems

2024-03-06 · Wesley A. Suttle, Vipul K. Sharma, Krishna C. Kosaraju, S. Sivaranjani, Ji Liu, Vijay Gupta, Brian M. Sadler

We develop provably safe and convergent reinforcement learning (RL) algorithms for control of nonlinear dynamical systems, bridging the gap between the hard safety guarantees of control theory and the convergence guarantees of RL theory. Recent advances at the intersection of control and RL follow a two-stage, safety filter approach to enforcing hard safety constraints: model-free RL is used to learn a potentially unsafe controller, whose actions are projected onto safe sets prescribed, for example, by a control barrier function. Though safe, such approaches lose any convergence guarantees enjoyed by the underlying RL methods. In this paper, we develop a single-stage, sampling-based approach to hard constraint satisfaction that learns RL controllers enjoying classical convergence guarantees while satisfying hard safety constraints throughout training and deployment. We validate the efficacy of our approach in simulation, including safe control of a quadcopter in a challenging obstacle avoidance problem, and demonstrate that it outperforms existing benchmarks.

📄 PDF Abstract BibTeX arXiv:2403.04007

Code (1)

sharma1256/cbf-constrained_ppo 공식 구현 pytorch

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Safe Reinforcement Learning

Similar Papers 제목 키워드 기반

Data Generation Method for Learning a Low-dimensional Safe Region in Safe Reinforcement Learning

2021-09-10 · Zhehua Zhou, Ozgur S. Oguz, Yi Ren, Marion Leibold 외

Safe reinforcement learning aims to learn a control policy while ensuring that neither the system nor the environment gets damaged during the learning process. For implementing safe reinforcement learning on highly nonli…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Safe Reinforcement Learning

A predictive safety filter for learning-based control of constrained nonlinear dynamical systems

2018-12-13 · Kim P. Wabersich, Melanie N. Zeilinger

The transfer of reinforcement learning (RL) techniques into real-world applications is challenged by safety requirements in the presence of physical limitations. Most RL methods, in particular the most popular algorithms…

Model Predictive ControlReinforcement LearningReinforcement Learning (RL)Safe Exploration

Learning Control Policies to Provably Satisfy Hard Affine Constraints for Black-Box Hybrid Dynamical Systems

2026-04-24 · Aayushi Shrivastava, Kartik Nagpal, Sairam Jinkala, Jean-Baptiste Bouvier 외 arxiv

Ensuring safety for black-box hybrid dynamical systems presents significant challenges due to their instantaneous state jumps and unknown explicit nonlinear dynamics. Existing solutions for strict safety constraint satis…

Reinforcement Learning

KCRL: Krasovskii-Constrained Reinforcement Learning with Guaranteed Stability in Nonlinear Dynamical Systems

2022-06-03 · Sahin Lale, Yuanyuan Shi, Guannan Qu, Kamyar Azizzadenesheli 외

Learning a dynamical system requires stabilizing the unknown dynamics to avoid state blow-ups. However, current reinforcement learning (RL) methods lack stabilization guarantees, which limits their applicability for the …

reinforcement-learningReinforcement Learning (RL)

Robust Model Predictive Shielding for Safe Reinforcement Learning with Stochastic Dynamics

2019-10-24 · Shuo Li, Osbert Bastani

This paper proposes a framework for safe reinforcement learning that can handle stochastic nonlinear dynamical systems. We focus on the setting where the nominal dynamics are known, and are subject to additive stochastic…

Learning Theoryreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1