paper-with-me

홈 › Papers

Inverse Transition Learning: Learning Dynamics from Demonstrations

2024-11-07 · Leo Benac, Abhishek Sharma, Sonali Parbhoo, Finale Doshi-Velez

We consider the problem of estimating the transition dynamics $T^*$ from near-optimal expert trajectories in the context of offline model-based reinforcement learning. We develop a novel constraint-based method, Inverse Transition Learning, that treats the limited coverage of the expert trajectories as a \emph{feature}: we use the fact that the expert is near-optimal to inform our estimate of $T^*$. We integrate our constraints into a Bayesian approach. Across both synthetic environments and real healthcare scenarios like Intensive Care Unit (ICU) patient management in hypotension, we demonstrate not only significant improvements in decision-making, but that our posterior can inform when transfer will be successful.

📄 PDF Abstract BibTeX arXiv:2411.05174

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingManagementModel-based Reinforcement Learning

Similar Papers 제목 키워드 기반

HMPC-assisted Adversarial Inverse Reinforcement Learning for Smart Home Energy Management

2025-06-01 · Jiadong He, Liang Yu, Zhiqiang Chen, Dawei Qiu 외

This letter proposes an Adversarial Inverse Reinforcement Learning (AIRL)-based energy management method for a smart home, which incorporates an implicit thermal dynamics model. In the proposed method, historical optimal…

energy managementManagementModel Predictive Controlreinforcement-learning+1

Inverse Reinforcement Learning with Simultaneous Estimation of Rewards and Dynamics

2016-04-13 · Michael Herman, Tobias Gindele, Jörg Wagner, Felix Schmitt 외

Inverse Reinforcement Learning (IRL) describes the problem of learning an unknown reward function of a Markov Decision Process (MDP) from observed behavior of an agent. Since the agent's behavior originates in its policy…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Transfer Learning

When Does Predictive Inverse Dynamics Outperform Behavior Cloning?

2026-01-29 · Lukas Schäfer, Pallavi Choudhury, Abdelhak Lemkhenter, Chris Lovett 외 arxiv

Behavior cloning (BC) is a practical offline imitation learning method, but it often fails when expert demonstrations are limited. Recent works have introduced a class of architectures named predictive inverse dynamics m…

Zero-shot Imitation Learning from Demonstrations for Legged Robot Visual Navigation

2019-09-27 · Xinlei Pan, Tingnan Zhang, Brian Ichter, Aleksandra Faust 외

Imitation learning is a popular approach for training visual navigation policies. However, collecting expert demonstrations for legged robots is challenging as these robots can be hard to control, move slowly, and cannot…

DisentanglementImitation LearningVisual Navigation

Model-Based Inverse Reinforcement Learning from Visual Demonstrations

2020-10-18 · Neha Das, Sarah Bechtle, Todor Davchev, Dinesh Jayaraman 외

Scaling model-based inverse reinforcement learning (IRL) to real robotic manipulation tasks with unknown dynamics remains an open problem. The key challenges lie in learning good dynamics models, developing algorithms th…

modelModel Predictive Controlreinforcement-learningReinforcement Learning+1