paper-with-me

홈 › Papers

Reinforcement Learning for Temporal Logic Control Synthesis with Probabilistic Satisfaction Guarantees

2019-09-11 · Mohammadhosein Hasanbeig, Yiannis Kantaros, Alessandro Abate, Daniel Kroening, George J. Pappas, Insup Lee

Reinforcement Learning (RL) has emerged as an efficient method of choice for solving complex sequential decision making problems in automatic control, computer science, economics, and biology. In this paper we present a model-free RL algorithm to synthesize control policies that maximize the probability of satisfying high-level control objectives given as Linear Temporal Logic (LTL) formulas. Uncertainty is considered in the workspace properties, the structure of the workspace, and the agent actions, giving rise to a Probabilistically-Labeled Markov Decision Process (PL-MDP) with unknown graph structure and stochastic behaviour, which is even more general case than a fully unknown MDP. We first translate the LTL specification into a Limit Deterministic Buchi Automaton (LDBA), which is then used in an on-the-fly product with the PL-MDP. Thereafter, we define a synchronous reward function based on the acceptance condition of the LDBA. Finally, we show that the RL algorithm delivers a policy that maximizes the satisfaction probability asymptotically. We provide experimental results that showcase the efficiency of the proposed method.

📄 PDF Abstract BibTeX arXiv:1909.05304

Code (1)

grockious/lcrl 공식 구현

Tasks

Decision MakingDecision Making Under UncertaintyHierarchical Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Safe Reinforcement LearningSequential Decision Making

Similar Papers 제목 키워드 기반

Reinforcement Learning Based Temporal Logic Control with Maximum Probabilistic Satisfaction

2020-10-14 · Mingyu Cai, Shaoping Xiao, Baoluo Li, Zhiliang Li 외

This paper presents a model-free reinforcement learning (RL) algorithm to synthesize a control policy that maximizes the satisfaction probability of linear temporal logic (LTL) specifications. Due to the consideration of…

Motion Planningreinforcement-learningReinforcement Learning (RL)

Provably Correct Controller Synthesis of Switched Stochastic Systems with Metric Temporal Logic Specifications: A Case Study on Power Systems

2021-03-26 · Zhe Xu, Yichen Zhang

In this paper, we present a provably correct controller synthesis approach for switched stochastic control systems with metric temporal logic (MTL) specifications with provable probabilistic guarantees. We first present …

Synthesis of Safety Specifications for Probabilistic Systems

2025-11-20 · Gaspard Ohlmann, Edwin Hamel-De le Court, Francesco Belardinelli arxiv

Ensuring that agents satisfy safety specifications can be crucial in safety-critical environments. While methods exist for controller synthesis with safe temporal specifications, most existing methods restrict safe tempo…

Synthesis of Provably Correct Autonomy Protocols for Shared Control

2019-05-15 · Murat Cubuktepe, Nils Jansen, Mohammed Alsiekh, Ufuk Topcu

We synthesize shared control protocols subject to probabilistic temporal logic specifications. More specifically, we develop a framework in which a human and an autonomy protocol can issue commands to carry out a certain…

Reinforcement Learning

Automaton-Guided Control Synthesis for Signal Temporal Logic Specifications

2022-07-08 · Qi Heng Ho, Roland B. Ilyes, Zachary N. Sunberg, Morteza Lahijanian

This paper presents an algorithmic framework for control synthesis of continuous dynamical systems subject to signal temporal logic (STL) specifications. We propose a novel algorithm to obtain a time-partitioned finite a…