paper-with-me

홈 › Papers

Reward Machines for Deep RL in Noisy and Uncertain Environments

2024-05-31 · Andrew C. Li, Zizhao Chen, Toryn Q. Klassen, Pashootan Vaezipoor, Rodrigo Toro Icarte, Sheila A. McIlraith

Reward Machines provide an automaton-inspired structure for specifying instructions, safety constraints, and other temporally extended reward-worthy behaviour. By exposing the underlying structure of a reward function, they enable the decomposition of an RL task, leading to impressive gains in sample efficiency. Although Reward Machines and similar formal specifications have a rich history of application towards sequential decision-making problems, they critically rely on a ground-truth interpretation of the domain-specific vocabulary that forms the building blocks of the reward function--such ground-truth interpretations are elusive in the real world due in part to partial observability and noisy sensing. In this work, we explore the use of Reward Machines for Deep RL in noisy and uncertain environments. We characterize this problem as a POMDP and propose a suite of RL algorithms that exploit task structure under uncertain interpretation of the domain-specific vocabulary. Through theory and experiments, we expose pitfalls in naive approaches to this problem while simultaneously demonstrating how task structure can be successfully leveraged under noisy interpretations of the vocabulary.

📄 PDF Abstract BibTeX arXiv:2406.00120

Code (1)

andrewli77/reward-machines-noisy-environments 공식 구현 pytorch

Tasks

counterfactualDecision MakingSequential Decision Making

Similar Papers 제목 키워드 기반

Noisy Symbolic Abstractions for Deep RL: A case study with Reward Machines

2022-11-20 · Andrew C. Li, Zizhao Chen, Pashootan Vaezipoor, Toryn Q. Klassen 외

Natural and formal languages provide an effective mechanism for humans to specify instructions and reward functions. We investigate how to generate policies via RL when reward functions are specified in a symbolic langua…

Assessing the Robustness of Intelligence-Driven Reinforcement Learning

2023-11-15 · Lorenzo Nodari, Federico Cerutti

Robustness to noise is of utmost importance in reinforcement learning systems, particularly in military contexts where high stakes and uncertain environments prevail. Noise and uncertainty are inherent features of milita…

Decision Makingreinforcement-learningReinforcement Learning

Reinforcement Learning with Stochastic Reward Machines

2025-10-16 · Jan Corazza, Ivan Gavran, Daniel Neider arxiv

Reward machines are an established tool for dealing with reinforcement learning problems in which rewards are sparse and depend on complex sequences of actions. However, existing algorithms for learning reward machines a…

Reinforcement Learning

Joint Learning of Reward Machines and Policies in Environments with Partially Known Semantics

2022-04-20 · Christos Verginis, Cevahir Koprulu, Sandeep Chinchali, Ufuk Topcu

We study the problem of reinforcement learning for a task encoded by a reward machine. The task is defined over a set of properties in the environment, called atomic propositions, and represented by Boolean variables. On…

Q-Learningreinforcement-learningReinforcement Learning (RL)

Learning Robust Reward Machines from Noisy Labels

2024-08-27 · Roko Parac, Lorenzo Nodari, Leo Ardon, Daniel Furelos-Blanco 외

This paper presents PROB-IRM, an approach that learns robust reward machines (RMs) for reinforcement learning (RL) agents from noisy execution traces. The key aspect of RM-driven RL is the exploitation of a finite-state …

Inductive logic programmingReinforcement Learning (RL)