paper-with-me

홈 › Papers

Offline Imitation Learning with a Misspecified Simulator

2020-12-01 · NeurIPS 2020 12 · Shengyi Jiang, JingCheng Pang, Yang Yu

In real-world decision-making tasks, learning an optimal policy without a trial-and-error process is an appealing challenge. When expert demonstrations are available, imitation learning that mimics expert actions can learn a good policy efficiently. Learning in simulators is another commonly adopted approach to avoid real-world trials-and-errors. However, neither sufficient expert demonstrations nor high-fidelity simulators are easy to obtain. In this work, we investigate policy learning in the condition of a few expert demonstrations and a simulator with misspecified dynamics. Under a mild assumption that local states shall still be partially aligned under a dynamics mismatch, we propose imitation learning with horizon-adaptive inverse dynamics (HIDIL) that matches the simulator states with expert states in a $H$-step horizon and accurately recovers actions based on inverse dynamics policies. In the real environment, HIDIL can effectively derive adapted actions from the matched states. Experiments are conducted in four MuJoCo locomotion environments with modified friction, gravity, and density configurations. Experiment results show that HIDIL achieves significant improvement in terms of performance and stability in all of the real environments, compared with imitation learning methods and transferring methods in reinforcement learning.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingFrictionImitation LearningMuJoCo

Similar Papers 제목 키워드 기반

Generalized Bayesian Inference for Scientific Simulators via Amortized Cost Estimation

2023-05-24 · NeurIPS 2023 11 · Richard Gao, Michael Deistler, Jakob H. Macke

Simulation-based inference (SBI) enables amortized Bayesian inference for simulators with implicit likelihoods. But when we are primarily interested in the quality of predictive simulations, or when the model cannot exac…

Bayesian Inference

When Does Online Imitation Learning Help in LLM Post-Training? The Role of (Non-)Realizability Beyond Horizon

2026-06-29 · Huaqing Zhang, Jingchu Gai, Juno Kim, Bingbin Liu 외 arxiv

Online imitation learning (IL), particularly on-policy distillation, has emerged as a strong LLM post-training approach, often outperforming offline supervised fine-tuning (SFT). Yet a principled understanding of when an…

A Recipe for Efficient Sim-to-Real Transfer in Manipulation with Online Imitation-Pretrained World Models

2025-10-02 · Yilin Wang, Shangzhe Li, Haoyi Niu, Zhiao Huang 외 arxiv

We are interested in solving the problem of imitation learning with a limited amount of real-world expert data. Existing offline imitation methods often struggle with poor data coverage and severe performance degradation…

Efficient Imitation under Misspecification

2025-03-17 · Nicolas Espinosa-Dice, Sanjiban Choudhury, Wen Sun, Gokul Swamy

We consider the problem of imitation learning under misspecification: settings where the learner is fundamentally unable to replicate expert behavior everywhere. This is often true in practice due to differences in obser…

Imitation Learning

Addressing Misspecification in Simulation-based Inference through Data-driven Calibration

2024-05-14 · Antoine Wehenkel, Juan L. Gamella, Ozan Sener, Jens Behrmann 외

Driven by steady progress in deep generative modeling, simulation-based inference (SBI) has emerged as the workhorse for inferring the parameters of stochastic simulators. However, recent work has demonstrated that model…