paper-with-me

홈 › Papers

Omega-Regular Objectives in Model-Free Reinforcement Learning

2018-09-26 · Hahn Ernst Moritz, Perez Mateo, Schewe Sven, Somenzi Fabio, Trivedi Ashutosh, Wojtczak Dominik

We provide the first solution for model-free reinforcement learning of {\omega}-regular objectives for Markov decision processes (MDPs). We present a constructive reduction from the almost-sure satisfaction of {\omega}-regular objectives to an almost- sure reachability problem and extend this technique to learning how to control an unknown model so that the chance of satisfying the objective is maximized. A key feature of our technique is the compilation of {\omega}-regular properties into limit- deterministic Buechi automata instead of the traditional Rabin automata; this choice sidesteps difficulties that have marred previous proposals. Our approach allows us to apply model-free, off-the-shelf reinforcement learning algorithms to compute optimal strategies from the observations of the MDP. We present an experimental evaluation of our technique on benchmark learning problems.

📄 PDF Abstract BibTeX arXiv:1810.00950

Code (0)

등록된 구현이 없습니다.

Tasks

modelreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Omega-Regular Reward Machines

2023-08-14 · Ernst Moritz Hahn, Mateo Perez, Sven Schewe, Fabio Somenzi 외

Reinforcement learning (RL) is a powerful approach for training agents to perform tasks, but designing an appropriate reward mechanism is critical to its success. However, in many cases, the complexity of the learning ob…

Reinforcement Learning (RL)

Reinforcement Learning for Omega-Regular Specifications on Continuous-Time MDP

2023-03-16 · Amin Falah, Shibashis Guha, Ashutosh Trivedi

Continuous-time Markov decision processes (CTMDPs) are canonical models to express sequential decision-making under dense-time and stochastic environments. When the stochastic evolution of the environment is only availab…

Decision Makingreinforcement-learningReinforcement LearningReinforcement Learning (RL)+2

A PAC Learning Algorithm for LTL and Omega-regular Objectives in MDPs

2023-10-18 · Mateo Perez, Fabio Somenzi, Ashutosh Trivedi

Linear temporal logic (LTL) and omega-regular objectives -- a superset of LTL -- have seen recent use as a way to express non-Markovian objectives in reinforcement learning. We introduce a model-based probably approximat…

PAC learningreinforcement-learning

Reinforcement Learning with LTL and $ω$-Regular Objectives via Optimality-Preserving Translation to Average Rewards

2024-10-16 · Xuan-Bach Le, Dominik Wagner, Leon Witzman, Alexander Rabinovich 외

Linear temporal logic (LTL) and, more generally, $\omega$-regular objectives are alternatives to the traditional discount sum and average reward objectives in reinforcement learning (RL), offering the advantage of greate…

Reinforcement Learning (RL)

Average Reward Reinforcement Learning for Omega-Regular and Mean-Payoff Objectives

2025-05-21 · Milad Kazemi, Mateo Perez, Fabio Somenzi, Sadegh Soudjani 외

Recent advances in reinforcement learning (RL) have renewed focus on the design of reward functions that shape agent behavior. Manually designing reward functions is tedious and error-prone. A principled alternative is t…

Reinforcement Learning (RL)