There Is No Turning Back: A Self-Supervised Approach for Reversibility-Aware Reinforcement Learning
We propose to learn to distinguish reversible from irreversible actions for better informed decision-making in Reinforcement Learning (RL). From theoretical considerations, we show that approximate reversibility can be learned through a simple surrogate task: ranking randomly sampled trajectory events in chronological order. Intuitively, pairs of events that are always observed in the same order are likely to be separated by an irreversible sequence of actions. Conveniently, learning the temporal order of events can be done in a fully self-supervised way, which we use to estimate the reversibility of actions from experience, without any priors. We propose two different strategies that incorporate reversibility in RL agents, one strategy for exploration (RAE) and one strategy for control (RAC). We demonstrate the potential of reversibility-aware agents in several environments, including the challenging Sokoban game. In synthetic tasks, we show that we can learn control policies that never fail and reduce to zero the side-effects of interactions, even without access to the reward function.
Code (0)
등록된 구현이 없습니다.
Tasks
Decision MakingReinforcement Learning (RL)SokobanSimilar Papers 제목 키워드 기반
Learning to Undo: Rollback-Augmented Reinforcement Learning with Reversibility Signals
This paper proposes a reversible learning framework to improve the robustness and efficiency of value based Reinforcement Learning agents, addressing vulnerability to value overestimation and instability in partially irr…
Reinforcement LearningDecision MakingDetermining ActionReversibility in STRIPS Using Answer Set and Epistemic Logic Programming
In the context of planning and reasoning about actions and change, we call an action reversible when its effects can be reverted by applying other actions, returning to the original state. Renewed interest in this area h…
TranslationTrend patterns statistics for assessing irreversibility in cryptocurrencies: time-asymmetry versus inefficiency
In this paper, we present a measure of time irreversibility using trend pattern statistics. We define the irreversibility index as the Kullback-Leibler divergence between the distribution of uptrends subsequences (increa…
Time SeriesRevisable by Design: A Theory of Streaming LLM Agent Execution
Current LLM agents operate under an implicit but universal assumption: execution is a transaction -- the user submits a request, the agent works in isolation, and only upon completion does the dialogue resume. This force…
Thermodynamic Irreversibility of Training Algorithms
The training algorithms for AI systems all introduce far-from-equilibrium dynamical processes, and understanding the irreversibility of these algorithms is a fundamental step towards understanding the learning dynamics o…