paper-with-me

홈 › Papers

Mungojerrie: Reinforcement Learning of Linear-Time Objectives

2021-06-16 · Ernst Moritz Hahn, Mateo Perez, Sven Schewe, Fabio Somenzi, Ashutosh Trivedi, Dominik Wojtczak

Reinforcement learning synthesizes controllers without prior knowledge of the system. At each timestep, a reward is given. The controllers optimize the discounted sum of these rewards. Applying this class of algorithms requires designing a reward scheme, which is typically done manually. The designer must ensure that their intent is accurately captured. This may not be trivial, and is prone to error. An alternative to this manual programming, akin to programming directly in assembly, is to specify the objective in a formal language and have it "compiled" to a reward scheme. Mungojerrie (https://plv.colorado.edu/mungojerrie/) is a tool for testing reward schemes for $\omega$-regular objectives on finite models. The tool contains reinforcement learning algorithms and a probabilistic model checker. Mungojerrie supports models specified in PRISM and $\omega$-automata specified in HOA.

📄 PDF Abstract BibTeX arXiv:2106.09161

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Multi-Objective Deep Reinforcement Learning

2016-10-09 · Hossam Mossalam, Yannis M. Assael, Diederik M. Roijers, Shimon Whiteson

We propose Deep Optimistic Linear Support Learning (DOL) to solve high-dimensional multi-objective decision problems where the relative importances of the objectives are not known a priori. Using features from the high-d…

Deep Reinforcement LearningMulti-Objective Reinforcement Learningreinforcement-learningReinforcement Learning+1

A PAC Learning Algorithm for LTL and Omega-regular Objectives in MDPs

2023-10-18 · Mateo Perez, Fabio Somenzi, Ashutosh Trivedi

Linear temporal logic (LTL) and omega-regular objectives -- a superset of LTL -- have seen recent use as a way to express non-Markovian objectives in reinforcement learning. We introduce a model-based probably approximat…

PAC learningreinforcement-learning

gTLO: A Generalized and Non-linear Multi-Objective Deep Reinforcement Learning Approach

2022-04-11 · Johannes Dornheim

In real-world decision optimization, often multiple competing objectives must be taken into account. Following classical reinforcement learning, these objectives have to be combined into a single reward function. In cont…

Deep Reinforcement LearningDeep-Sea Treasure, Image versionMulti-Objective Reinforcement Learningreinforcement-learning+2

Joint Optimization of Multi-Objective Reinforcement Learning with Policy Gradient Based Algorithm

2021-05-28 · Qinbo Bai, Mridul Agarwal, Vaneet Aggarwal

Many engineering problems have multiple objectives, and the overall aim is to optimize a non-linear function of these objectives. In this paper, we formulate the problem of maximizing a non-linear concave function of mul…

Multi-Objective Reinforcement Learningreinforcement-learningReinforcement Learning (RL)

Computably Continuous Reinforcement-Learning Objectives are PAC-learnable

2023-03-09 · Cambridge Yang, Michael Littman, Michael Carbin

In reinforcement learning, the classic objectives of maximizing discounted and finite-horizon cumulative rewards are PAC-learnable: There are algorithms that learn a near-optimal policy with high probability using a fini…

General Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)