paper-with-me

홈 › Papers

Inverse Reinforcement Learning with Simultaneous Estimation of Rewards and Dynamics

2016-04-13 · Michael Herman, Tobias Gindele, Jörg Wagner, Felix Schmitt, Wolfram Burgard

Inverse Reinforcement Learning (IRL) describes the problem of learning an unknown reward function of a Markov Decision Process (MDP) from observed behavior of an agent. Since the agent's behavior originates in its policy and MDP policies depend on both the stochastic system dynamics as well as the reward function, the solution of the inverse problem is significantly influenced by both. Current IRL approaches assume that if the transition model is unknown, additional samples from the system's dynamics are accessible, or the observed behavior provides enough samples of the system's dynamics to solve the inverse problem accurately. These assumptions are often not satisfied. To overcome this, we present a gradient-based IRL approach that simultaneously estimates the system's dynamics. By solving the combined optimization problem, our approach takes into account the bias of the demonstrations, which stems from the generating policy. The evaluation on a synthetic MDP and a transfer learning task shows improvements regarding the sample efficiency as well as the accuracy of the estimated reward functions and transition models.

📄 PDF Abstract BibTeX arXiv:1604.03912

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Transfer Learning

Similar Papers 제목 키워드 기반

Adversarial Imitation via Variational Inverse Reinforcement Learning

2018-09-17 · ICLR 2019 5 · Ahmed H. Qureshi, Byron Boots, Michael C. Yip

We consider a problem of learning the reward and policy from expert examples under unknown dynamics. Our proposed method builds on the framework of generative adversarial networks and introduces the empowerment-regulariz…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Transfer Learning

A Bayesian Approach to Robust Inverse Reinforcement Learning

2023-09-15 · Ran Wei, Siliang Zeng, Chenliang Li, Alfredo Garcia 외

We consider a Bayesian approach to offline model-based inverse reinforcement learning (IRL). The proposed framework differs from existing offline model-based IRL approaches by performing simultaneous estimation of the ex…

Imitation LearningMuJoCoreinforcement-learningReinforcement Learning

oIRL: Robust Adversarial Inverse Reinforcement Learning with Temporally Extended Actions

2020-02-20 · David Venuto, Jhelum Chakravorty, Leonard Boussioux, Junhao Wang 외

Explicit engineering of reward functions for given environments has been a major hindrance to reinforcement learning methods. While Inverse Reinforcement Learning (IRL) is a solution to recover reward functions from demo…

continuous-controlContinuous Controlreinforcement-learningReinforcement Learning+2

Learning Robust Rewards with Adversarial Inverse Reinforcement Learning

2017-10-30 · Justin Fu, Katie Luo, Sergey Levine

Reinforcement learning provides a powerful and general framework for decision making and control, but its application in practice is often hindered by the need for extensive feature and reward engineering. Deep reinforce…

Decision MakingDeep Reinforcement Learningreinforcement-learningReinforcement Learning+1

Maximum-Likelihood Inverse Reinforcement Learning with Finite-Time Guarantees

2022-10-04 · Siliang Zeng, Chenliang Li, Alfredo Garcia, Mingyi Hong

Inverse reinforcement learning (IRL) aims to recover the reward function and the associated optimal policy that best fits observed sequences of states and actions implemented by an expert. Many algorithms for IRL have an…

counterfactualImitation LearningMuJoCoreinforcement-learning+2