paper-with-me

Papers

Learning Time-Invariant Reward Functions through Model-Based Inverse Reinforcement Learning

2021-07-07 · Todor Davchev, Sarah Bechtle, Subramanian Ramamoorthy, Franziska Meier

Inverse reinforcement learning is a paradigm motivated by the goal of learning general reward functions from demonstrated behaviours. Yet the notion of generality for learnt costs is often evaluated in terms of robustness to various spatial perturbations only, assuming deployment at fixed speeds of execution. However, this is impractical in the context of robotics and building, time-invariant solutions is of crucial importance. In this work, we propose a formulation that allows us to 1) vary the length of execution by learning time-invariant costs, and 2) relax the temporal alignment requirements for learning from demonstration. We apply our method to two different types of cost formulations and evaluate their performance in the context of learning reward functions for simulated placement and peg in hole tasks executed on a 7DoF Kuka IIWA arm. Our results show that our approach enables learning temporally invariant rewards from misaligned demonstration that can also generalise spatially to out of distribution tasks.

📄 PDF Abstract BibTeX arXiv:2107.03186

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Learning Invariant Reward Functions through Trajectory Interventions

2021-09-29 · Ivan Ovinnikov, Eugene Bykovets, Joachim M. Buhmann

Inverse reinforcement learning methods aim to retrieve the reward function of a Markov decision process based on a dataset of expert demonstrations. The commonplace scarcity of such demonstrations potentially leads to th…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Statistical analysis of Inverse Entropy-regularized Reinforcement Learning

2025-12-07 · Denis Belomestny, Alexey Naumov, Sergey Samsonov arxiv

Inverse reinforcement learning aims to infer the reward function that explains expert behavior observed through trajectories of state--action pairs. A long-standing difficulty in classical IRL is the non-uniqueness of th…

Reinforcement Learning

Learning Causally Invariant Reward Functions from Diverse Demonstrations

2024-09-12 · Ivan Ovinnikov, Eugene Bykovets, Joachim M. Buhmann

Inverse reinforcement learning methods aim to retrieve the reward function of a Markov decision process based on a dataset of expert demonstrations. The commonplace scarcity and heterogeneous sources of such demonstratio…

reinforcement-learningReinforcement Learning

Efficient Exploration of Reward Functions in Inverse Reinforcement Learning via Bayesian Optimization

2020-11-17 · NeurIPS 2020 12 · Sreejith Balakrishnan, Quoc Phong Nguyen, Bryan Kian Hsiang Low, Harold Soh

The problem of inverse reinforcement learning (IRL) is relevant to a variety of tasks including value alignment and robot learning from demonstration. Despite significant algorithmic contributions in recent years, IRL re…

Bayesian OptimizationEfficient Explorationreinforcement-learningReinforcement Learning (RL)

Adversarial recovery of agent rewards from latent spaces of the limit order book

2019-12-09 · Jacobo Roa-Vicens, Yuanbo Wang, Virgile Mison, Yarin Gal 외

Inverse reinforcement learning has proved its ability to explain state-action trajectories of expert agents by recovering their underlying reward functions in increasingly challenging environments. Recent advances in adv…

Reinforcement Learning