paper-with-me

홈 › Papers

On exploration requirements for learning safety constraints

2021-05-17 · Pierre-François Massiani, Steve Heim, Sebastian Trimpe

Enforcing safety for dynamical systems is challenging, since it requires constraint satisfaction along trajectory predictions. Equivalent control constraints can be computed in the form of sets that enforce positive invariance, and can thus guarantee safety in feedback controllers without predictions. However, these constraints are cumbersome to compute from models, and it is not yet well established how to infer constraints from data. In this paper, we shed light on the key objects involved in learning control constraints from data in a model-free setting. In particular, we discuss the family of constraints that enforce safety in the context of a nominal control policy, and expose that these constraints do not need to be accurate everywhere. They only need to correctly exclude a subset of the state-actions that would cause failure, which we call the critical set.

📄 PDF Abstract BibTeX arXiv:2105.08143

Code (1)

Data-Science-in-Mechanical-Engineering/edge 공식 구현

Similar Papers 제목 키워드 기반

DOPE: Doubly Optimistic and Pessimistic Exploration for Safe Reinforcement Learning

2021-12-01 · Archana Bura, Aria HasanzadeZonuzy, Dileep Kalathil, Srinivas Shakkottai 외

Safe reinforcement learning is extremely challenging--not only must the agent explore an unknown environment, it must do so while ensuring no safety constraint violations. We formulate this safe reinforcement learning (R…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Safe Exploration+1

Safe and Adaptive Decision-Making for Optimization of Safety-Critical Systems: The ARTEO Algorithm

2022-11-10 · Buse Sibel Korkmaz, Marta Zagórowska, Mehmet Mercangöz

We consider the problem of decision-making under uncertainty in an environment with safety constraints. Many business and industrial applications rely on real-time optimization to improve key performance indicators. In t…

Decision MakingDecision Making Under UncertaintyGaussian ProcessesMulti-Armed Bandits

Efficiently Computable Safety Bounds for Gaussian Processes in Active Learning

2024-02-28 · Jörn Tebbe, Christoph Zimmer, Ansgar Steland, Markus Lange-Hegermann 외

Active learning of physical systems must commonly respect practical safety constraints, which restricts the exploration of the design space. Gaussian Processes (GPs) and their calibrated uncertainty estimations are widel…

Active LearningGaussian Processes

Handling Long-Term Safety and Uncertainty in Safe Reinforcement Learning

2024-09-18 · Jonas Günster, Puze Liu, Jan Peters, Davide Tateo

Safety is one of the key issues preventing the deployment of reinforcement learning techniques in real-world robots. While most approaches in the Safe Reinforcement Learning area do not require prior knowledge of constra…

reinforcement-learningReinforcement LearningSafe ExplorationSafe Reinforcement Learning

Constrained Exploration and Recovery from Experience Shaping

2018-09-21 · Tu-Hoa Pham, Giovanni De Magistris, Don Joven Agravante, Subhajit Chaudhury 외

We consider the problem of reinforcement learning under safety requirements, in which an agent is trained to complete a given task, typically formalized as the maximization of a reward signal over time, while concurrentl…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)