paper-with-me

홈 › Papers

There Is No Turning Back: A Self-Supervised Approach for Reversibility-Aware Reinforcement Learning

2021-06-08 · NeurIPS 2021 12 · Nathan Grinsztajn, Johan Ferret, Olivier Pietquin, Philippe Preux, Matthieu Geist

We propose to learn to distinguish reversible from irreversible actions for better informed decision-making in Reinforcement Learning (RL). From theoretical considerations, we show that approximate reversibility can be learned through a simple surrogate task: ranking randomly sampled trajectory events in chronological order. Intuitively, pairs of events that are always observed in the same order are likely to be separated by an irreversible sequence of actions. Conveniently, learning the temporal order of events can be done in a fully self-supervised way, which we use to estimate the reversibility of actions from experience, without any priors. We propose two different strategies that incorporate reversibility in RL agents, one strategy for exploration (RAE) and one strategy for control (RAC). We demonstrate the potential of reversibility-aware agents in several environments, including the challenging Sokoban game. In synthetic tasks, we show that we can learn control policies that never fail and reduce to zero the side-effects of interactions, even without access to the reward function.

📄 PDF Abstract BibTeX arXiv:2106.04480

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingReinforcement Learning (RL)Sokoban

Similar Papers 제목 키워드 기반

Learning to Undo: Rollback-Augmented Reinforcement Learning with Reversibility Signals

2025-10-16 · Andrejs Sorstkins, Omer Tariq, Muhammad Bilal arxiv

This paper proposes a reversible learning framework to improve the robustness and efficiency of value based Reinforcement Learning agents, addressing vulnerability to value overestimation and instability in partially irr…

Reinforcement LearningDecision Making

Determining ActionReversibility in STRIPS Using Answer Set and Epistemic Logic Programming

2021-08-11 · Wolfgang Faber, Michael Morak, Lukáš Chrpa

In the context of planning and reasoning about actions and change, we call an action reversible when its effects can be reverted by applying other actions, returning to the original state. Renewed interest in this area h…

Translation

Trend patterns statistics for assessing irreversibility in cryptocurrencies: time-asymmetry versus inefficiency

2023-06-28 · Jessica Morales Herrera, Raúl Salgado-García

In this paper, we present a measure of time irreversibility using trend pattern statistics. We define the irreversibility index as the Kullback-Leibler divergence between the distribution of uptrends subsequences (increa…

Time Series

Revisable by Design: A Theory of Streaming LLM Agent Execution

2026-04-25 · Zhiyuan Zhai, Ming Li, Xin Wang arxiv

Current LLM agents operate under an implicit but universal assumption: execution is a transaction -- the user submits a request, the agent works in isolation, and only upon completion does the dialogue resume. This force…

Thermodynamic Irreversibility of Training Algorithms

2026-05-21 · Liu Ziyin, Yuanjie Ren, Adam Levine, Isaac Chuang arxiv

The training algorithms for AI systems all introduce far-from-equilibrium dynamical processes, and understanding the irreversibility of these algorithms is a fundamental step towards understanding the learning dynamics o…