paper-with-me

홈 › Papers

Value Functions as Supermartingale Certificates

2026-05-29 · Alessandro Abate, Daniel Contro, Mirco Giacobbe, Agustín Martínez-Suñé, Diptarko Roy arxiv

Certification methods for stochastic systems provide sufficient proof rules, based on real-valued supermartingale certificates, to determine the almost-sure satisfaction of $ω$-regular properties (and therefore of linear temporal logic) over general state spaces, encompassing both countably infinite and continuous state spaces. Conversely, reinforcement learning (RL) methods for $ω$-regular tasks have received considerable attention, but they typically lack formal guarantees that the learned policy satisfies the specification, except possibly for finite state and action spaces. We bridge these two lines of research by establishing a novel theoretical connection: under an appropriate reward, the value function associated to a policy that almost surely satisfies an $ω$-regular property encodes a Streett supermartingale certificate for that specification. Our results, validated experimentally on finite Markov decision processes, hold for finite, countably infinite, and continuous state spaces, suggesting a principled route to certificate synthesis via RL.

📄 PDF Abstract BibTeX arXiv:2605.31524

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Neural Continuous-Time Supermartingale Certificates

2024-12-23 · Grigory Neustroev, Mirco Giacobbe, Anna Lukina

We introduce for the first time a neural-certificate framework for continuous-time stochastic dynamical systems. Autonomous learning systems in the physical world demand continuous-time reasoning, yet existing learnable …

Safety Certification for Stochastic Systems via Neural Barrier Functions

2022-06-03 · Frederik Baymler Mathiesen, Simeon Calvert, Luca Laurenti

Providing non-trivial certificates of safety for non-linear stochastic systems is an important open problem that limits the wider adoption of autonomous systems in safety-critical applications. One promising solution to …

A Supermartingale Relation for Multivariate Risk Measures

2018-01-31

The equivalence between multiportfolio time consistency of a dynamic multivariate risk measure and a supermartingale property is proven. Furthermore, the dual variables under which this set-valued supermartingale is a ma…

Relation

Learning Control Policies for Stochastic Systems with Reach-avoid Guarantees

2022-10-11 · Đorđe Žikelić, Mathias Lechner, Thomas A. Henzinger, Krishnendu Chatterjee

We study the problem of learning controllers for discrete-time non-linear stochastic dynamical systems with formal reach-avoid guarantees. This work presents the first method for providing formal reach-avoid guarantees, …

Policy Verification in Stochastic Dynamical Systems Using Logarithmic Neural Certificates

2024-06-02 · Thom Badings, Wietze Koops, Sebastian Junges, Nils Jansen

We consider the verification of neural network policies for discrete-time stochastic systems with respect to reach-avoid specifications. We use a learner-verifier procedure that learns a certificate for the specification…