paper-with-me

홈 › Papers

Learning to Provably Satisfy High Relative Degree Constraints for Black-Box Systems

2024-07-29 · Jean-Baptiste Bouvier, Kartik Nagpal, Negar Mehr

In this paper, we develop a method for learning a control policy guaranteed to satisfy an affine state constraint of high relative degree in closed loop with a black-box system. Previous reinforcement learning (RL) approaches to satisfy safety constraints either require access to the system model, or assume control affine dynamics, or only discourage violations with reward shaping. Only recently have these issues been addressed with POLICEd RL, which guarantees constraint satisfaction for black-box systems. However, this previous work can only enforce constraints of relative degree 1. To address this gap, we build a novel RL algorithm explicitly designed to enforce an affine state constraint of high relative degree in closed loop with a black-box control system. Our key insight is to make the learned policy be affine around the unsafe set and to use this affine region to dissipate the inertia of the high relative degree constraint. We prove that such policies guarantee constraint satisfaction for deterministic systems while being agnostic to the choice of the RL training algorithm. Our results demonstrate the capacity of our approach to enforce hard constraints in the Gym inverted pendulum and on a space shuttle landing simulation.

📄 PDF Abstract BibTeX arXiv:2407.20456

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Probabilistic Safety Constraints for Learned High Relative Degree System Dynamics

2019-12-20 · L4DC 2020 6 · Mohammad Javad Khojasteh, Vikas Dhiman, Massimo Franceschetti, Nikolay Atanasov

This paper focuses on learning a model of system dynamics online while satisfying safety constraints.Our motivation is to avoid offline system identification or hand-specified dynamics models and allowa system to safely …

Vocal Bursts Intensity Prediction

Rectified Control Barrier Functions for High-Order Safety Constraints

2024-12-04 · Pio Ong, Max H. Cohen, Tamas G. Molnar, Aaron D. Ames

This paper presents a novel approach for synthesizing control barrier functions (CBFs) from high relative degree safety constraints: Rectified CBFs (ReCBFs). We begin by discussing the limitations of existing High-Order …

Learning Control Policies to Provably Satisfy Hard Affine Constraints for Black-Box Hybrid Dynamical Systems

2026-04-24 · Aayushi Shrivastava, Kartik Nagpal, Sairam Jinkala, Jean-Baptiste Bouvier 외 arxiv

Ensuring safety for black-box hybrid dynamical systems presents significant challenges due to their instantaneous state jumps and unknown explicit nonlinear dynamics. Existing solutions for strict safety constraint satis…

Reinforcement Learning

Compatibility of Multiple Control Barrier Functions for Constrained Nonlinear Systems

2025-09-04 · Max H. Cohen, Eugene Lavretsky, Aaron D. Ames arxiv

Control barrier functions (CBFs) are a powerful tool for the constrained control of nonlinear systems; however, the majority of results in the literature focus on systems subject to a single CBF constraint, making it cha…

Small-Gain Theorem for Safety Verification under High-Relative-Degree Constraints

2022-04-09 · Ziliang Lyu, Xiangru Xu, Yiguang Hong

This paper develops a small-gain technique for the safety analysis and verification of interconnected systems with high-relative-degree safety constraints. In this technique, input-to-state safety (ISSf) is used to chara…

Vocal Bursts Intensity Prediction