paper-with-me

홈 › Papers

A Bayesian Approach to Robust Inverse Reinforcement Learning

2023-09-15 · Ran Wei, Siliang Zeng, Chenliang Li, Alfredo Garcia, Anthony McDonald, Mingyi Hong

We consider a Bayesian approach to offline model-based inverse reinforcement learning (IRL). The proposed framework differs from existing offline model-based IRL approaches by performing simultaneous estimation of the expert's reward function and subjective model of environment dynamics. We make use of a class of prior distributions which parameterizes how accurate the expert's model of the environment is to develop efficient algorithms to estimate the expert's reward and subjective dynamics in high-dimensional settings. Our analysis reveals a novel insight that the estimated policy exhibits robust performance when the expert is believed (a priori) to have a highly accurate model of the environment. We verify this observation in the MuJoCo environments and show that our algorithms outperform state-of-the-art offline IRL algorithms.

📄 PDF Abstract BibTeX arXiv:2309.08571

Code (2)

ran-weii/bmirl_tf 공식 구현 pytorch
ran-weii/cleanil pytorch

Tasks

Imitation LearningMuJoCoreinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

Multi-agent Inverse Reinforcement Learning for Two-person Zero-sum Games

2014-03-25 · Xiaomin Lin, Peter A. Beling, Randy Cogill

The focus of this paper is a Bayesian framework for solving a class of problems termed multi-agent inverse reinforcement learning (MIRL). Compared to the well-known inverse reinforcement learning (IRL) problem, MIRL is f…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Vocal Bursts Valence Prediction

Necessary and Sufficient Conditions for Inverse Reinforcement Learning of Bayesian Stopping Time Problems

2020-07-07 · Kunal Pattanayak, Vikram Krishnamurthy

This paper presents an inverse reinforcement learning~(IRL) framework for Bayesian stopping time problems. By observing the actions of a Bayesian decision maker, we provide a necessary and sufficient condition to identif…

reinforcement-learningReinforcement Learning (RL)Two-sample testing

Nonparametric Bayesian Inverse Reinforcement Learning for Multiple Reward Functions

2012-12-01 · NeurIPS 2012 12 · Jaedeug Choi, Kee-Eung Kim

We present a nonparametric Bayesian approach to inverse reinforcement learning (IRL) for multiple reward functions. Most previous IRL algorithms assume that the behaviour data is obtained from an agent who is optimizing …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization

2025-07-06 · Vikram Krishnamurthy arxiv

This monograph, spanning three chapters, explores Inverse Reinforcement Learning (IRL). The first two chapters view inverse reinforcement learning (IRL) through the lens of revealed preferences from microeconomics while …

Stochastic OptimizationReinforcement Learning

Kernel Density Bayesian Inverse Reinforcement Learning

2023-03-13 · Aishwarya Mandyam, Didong Li, Diana Cai, Andrew Jones 외

Inverse reinforcement learning (IRL) methods infer an agent's reward function using demonstrations of expert behavior. A Bayesian IRL approach models a distribution over candidate reward functions, capturing a degree of …

BIRLDensity Estimationreinforcement-learningReinforcement Learning+1