paper-with-me

홈 › Papers

Model-Based Reward Shaping for Adversarial Inverse Reinforcement Learning in Stochastic Environments

2024-10-04 · Simon Sinong Zhan, Qingyuan Wu, Philip Wang, YiXuan Wang, Ruochen Jiao, Chao Huang, Qi Zhu

In this paper, we aim to tackle the limitation of the Adversarial Inverse Reinforcement Learning (AIRL) method in stochastic environments where theoretical results cannot hold and performance is degraded. To address this issue, we propose a novel method which infuses the dynamics information into the reward shaping with the theoretical guarantee for the induced optimal policy in the stochastic environments. Incorporating our novel model-enhanced rewards, we present a novel Model-Enhanced AIRL framework, which integrates transition model estimation directly into reward shaping. Furthermore, we provide a comprehensive theoretical analysis of the reward error bound and performance difference bound for our method. The experimental results in MuJoCo benchmarks show that our method can achieve superior performance in stochastic environments and competitive performance in deterministic environments, with significant improvement in sample efficiency, compared to existing baselines.

📄 PDF Abstract BibTeX arXiv:2410.03847

Code (0)

등록된 구현이 없습니다.

Tasks

MuJoCo

Similar Papers 제목 키워드 기반

Toward Computationally Efficient Inverse Reinforcement Learning via Reward Shaping

2023-12-15 · Lauren H. Cooke, Harvey Klyne, Edwin Zhang, Cassidy Laidlaw 외

Inverse reinforcement learning (IRL) is computationally challenging, with common approaches requiring the solution of multiple reinforcement learning (RL) sub-problems. This work motivates the use of potential-based rewa…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Unbiased learning with State-Conditioned Rewards in Adversarial Imitation Learning

2021-01-01 · Dong-Sig Han, Hyunseo Kim, Hyundo Lee, Je-Hwan Ryu 외

Adversarial imitation learning has emerged as a general and scalable framework for automatic reward acquisition. However, we point out that previous methods commonly exploited occupancy-dependent reward learning formulat…

continuous-controlContinuous ControlImitation Learningreinforcement-learning+2

Competitive Multi-agent Inverse Reinforcement Learning with Sub-optimal Demonstrations

2018-01-07 · ICML 2018 7 · Xingyu Wang, Diego Klabjan

This paper considers the problem of inverse reinforcement learning in zero-sum stochastic games when expert demonstrations are known to be not optimal. Compared to previous works that decouple agents in the game by assum…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

IR$^3$: Contrastive Inverse Reinforcement Learning for Interpretable Detection and Mitigation of Reward Hacking

2026-02-23 · Mohammad Beigi, Ming Jin, Junshan Zhang, Jiaxin Zhang 외 arxiv

Reinforcement Learning from Human Feedback (RLHF) enables powerful LLM alignment but can introduce reward hacking - models exploit spurious correlations in proxy rewards without genuine alignment. Compounding this, the o…

Reinforcement Learning

Adversarial recovery of agent rewards from latent spaces of the limit order book

2019-12-09 · Jacobo Roa-Vicens, Yuanbo Wang, Virgile Mison, Yarin Gal 외

Inverse reinforcement learning has proved its ability to explain state-action trajectories of expert agents by recovering their underlying reward functions in increasingly challenging environments. Recent advances in adv…

Reinforcement Learning