paper-with-me

Papers

Learning Control Policies for Stochastic Systems with Reach-avoid Guarantees

2022-10-11 · Đorđe Žikelić, Mathias Lechner, Thomas A. Henzinger, Krishnendu Chatterjee

We study the problem of learning controllers for discrete-time non-linear stochastic dynamical systems with formal reach-avoid guarantees. This work presents the first method for providing formal reach-avoid guarantees, which combine and generalize stability and safety guarantees, with a tolerable probability threshold $p\in[0,1]$ over the infinite time horizon. Our method leverages advances in machine learning literature and it represents formal certificates as neural networks. In particular, we learn a certificate in the form of a reach-avoid supermartingale (RASM), a novel notion that we introduce in this work. Our RASMs provide reachability and avoidance guarantees by imposing constraints on what can be viewed as a stochastic extension of level sets of Lyapunov functions for deterministic systems. Our approach solves several important problems -- it can be used to learn a control policy from scratch, to verify a reach-avoid specification for a fixed control policy, or to fine-tune a pre-trained policy if it does not satisfy the reach-avoid specification. We validate our approach on $3$ stochastic non-linear reinforcement learning tasks.

📄 PDF Abstract BibTeX arXiv:2210.05308

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Compositional Policy Learning in Stochastic Control Systems with Formal Guarantees

2023-12-03 · NeurIPS 2023 11 · Đorđe Žikelić, Mathias Lechner, Abhinav Verma, Krishnendu Chatterjee 외

Reinforcement learning has shown promising results in learning neural network policies for complicated control tasks. However, the lack of formal guarantees about the behavior of such policies remains an impediment to th…

A learning-based approach to stochastic optimal control under reach-avoid constraint

2024-12-21 · Tingting Ni, Maryam Kamgarpour

We develop a model-free approach to optimally control stochastic, Markovian systems subject to a reach-avoid constraint. Specifically, the state trajectory must remain within a safe set while reaching a target set within…

Distributionally Robust Control Synthesis for Stochastic Systems with Safety and Reach-Avoid Specifications

2025-01-06 · Yu Chen, Yuda Li, ShaoYuan Li, Xiang Yin

We investigate the problem of synthesizing distributionally robust control policies for stochastic systems under safety and reach-avoid specifications. Using a game-theoretical framework, we consider the setting where th…

Probabilistic Reach-Avoid for Bayesian Neural Networks

2023-10-03 · Matthew Wicker, Luca Laurenti, Andrea Patane, Nicola Paoletti 외

Model-based reinforcement learning seeks to simultaneously learn the dynamics of an unknown stochastic environment and synthesise an optimal policy for acting in it. Ensuring the safety and robustness of sequential decis…

Model-based Reinforcement Learning

Markov Potential Game with Final-time Reach-Avoid Objectives

2024-10-23 · Sarah H. Q. Li, Abraham P. Vinod

We formulate a Markov potential game with final-time reach-avoid objectives by integrating potential game theory with stochastic reach-avoid control. Our focus is on multi-player trajectory planning where players maximiz…

Motion PlanningTrajectory Planning