paper-with-me

홈 › Papers

Repeated Inverse Reinforcement Learning

2017-05-15 · NeurIPS 2017 12 · Kareem Amin, Nan Jiang, Satinder Singh

We introduce a novel repeated Inverse Reinforcement Learning problem: the agent has to act on behalf of a human in a sequence of tasks and wishes to minimize the number of tasks that it surprises the human by acting suboptimally with respect to how the human would have acted. Each time the human is surprised, the agent is provided a demonstration of the desired behavior by the human. We formalize this problem, including how the sequence of tasks is chosen, in a few different ways and provide some foundational results.

📄 PDF Abstract BibTeX arXiv:1705.05427

Code (0)

등록된 구현이 없습니다.

Tasks

Imitation Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Inverse Reinforcement Learning without Reinforcement Learning

2023-03-26 · Gokul Swamy, Sanjiban Choudhury, J. Andrew Bagnell, Zhiwei Steven Wu

Inverse Reinforcement Learning (IRL) is a powerful set of techniques for imitation learning that aims to learn a reward function that rationalizes expert demonstrations. Unfortunately, traditional IRL methods suffer from…

continuous-controlContinuous ControlImitation Learningreinforcement-learning+2

Environment Design for Inverse Reinforcement Learning

2022-10-26 · Thomas Kleine Buening, Victor Villin, Christos Dimitrakakis

Learning a reward function from demonstrations suffers from low sample-efficiency. Even with abundant data, current inverse reinforcement learning methods that focus on learning from a single environment can fail to hand…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Hybrid Inverse Reinforcement Learning

2024-02-13 · Juntao Ren, Gokul Swamy, Zhiwei Steven Wu, J. Andrew Bagnell 외

The inverse reinforcement learning approach to imitation learning is a double-edged sword. On the one hand, it can enable learning from a smaller number of expert demonstrations with more robustness to error compounding …

continuous-controlContinuous ControlImitation Learningreinforcement-learning+2

TW-CRL: Time-Weighted Contrastive Reward Learning for Efficient Inverse Reinforcement Learning

2025-04-08 · YuXuan Li, Yicheng Gao, Ning Yang, Stephen Xia

Episodic tasks in Reinforcement Learning (RL) often pose challenges due to sparse reward signals and high-dimensional state spaces, which hinder efficient learning. Additionally, these tasks often feature hidden "trap st…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

The Virtues of Pessimism in Inverse Reinforcement Learning

2024-02-04 · David Wu, Gokul Swamy, J. Andrew Bagnell, Zhiwei Steven Wu 외

Inverse Reinforcement Learning (IRL) is a powerful framework for learning complex behaviors from expert demonstrations. However, it traditionally requires repeatedly solving a computationally expensive reinforcement lear…

Offline RLreinforcement-learningReinforcement LearningReinforcement Learning (RL)