paper-with-me

Papers

Model Predictive Adversarial Imitation Learning for Planning from Observation

2025-07-29 · Tyler Han, Yanda Bao, Bhaumik Mehta, Gabriel Guo, Anubhav Vishwakarma, Emily Kang, Sanghun Jung, Rosario Scalise, Jason Zhou, Bryan Xu, Byron Boots arxiv

Human demonstration data is often ambiguous and incomplete, motivating imitation learning approaches that also exhibit reliable planning behavior. A common paradigm to perform planning-from-demonstration involves learning a reward function via Inverse Reinforcement Learning (IRL) then deploying this reward via Model Predictive Control (MPC). Towards unifying these methods, we derive a replacement of the policy in IRL with a planning-based agent. With connections to Adversarial Imitation Learning, this formulation enables end-to-end interactive learning of planners from observation-only demonstrations. In addition to benefits in interpretability, complexity, and safety, we study and observe significant improvements on sample efficiency, out-of-distribution generalization, and robustness. The study includes evaluations in both simulated control benchmarks and real-world navigation experiments using few-to-single observation-only demonstrations.

📄 PDF Abstract BibTeX arXiv:2507.21533

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

CLAW: Learning Continuous Latent Action World Models via Adversarial Latent Regularization

2026-06-02 · Tewodros Ayalew, Matthew Jeung, Samuel Wheeler, Xiao Zhang 외 arxiv

We introduce CLAW, a fully end-to-end self-supervised framework for learning a world model jointly with continuous latent action representations directly from action-free videos. Our approach leverages adversarial latent…

Video Generation

Beyond Static Assumptions: the Predictive Justified Perspective Model for Epistemic Planning

2024-12-10 · Weijia Li, Guang Hu, Yangmengfei Xu

Epistemic Planning (EP) is an important research area dedicated to reasoning about the knowledge and beliefs of agents in multi-agent cooperative or adversarial settings. The Justified Perspective (JP) model is the state…

Understanding Adversarial Imitation Learning in Small Sample Regime: A Stage-coupled Analysis

2022-08-03 · Tian Xu, Ziniu Li, Yang Yu, Zhi-Quan Luo

Imitation learning learns a policy from expert trajectories. While the expert data is believed to be crucial for imitation quality, it was found that a kind of imitation learning approach, adversarial imitation learning …

Imitation Learning

Dyna-AIL : Adversarial Imitation Learning by Planning

2019-03-08 · Vaibhav Saxena, Srinivasan Sivanandan, Pulkit Mathur

Adversarial methods for imitation learning have been shown to perform well on various control tasks. However, they require a large number of environment interactions for convergence. In this paper, we propose an end-to-e…

Imitation Learning

NavForesee: A Unified Vision-Language World Model for Hierarchical Planning and Dual-Horizon Navigation Prediction

2025-12-01 · Fei Liu, Shichao Xie, Minghua Luo, Zedong Chu 외 arxiv

Embodied navigation for long-horizon tasks, guided by complex natural language instructions, remains a formidable challenge in artificial intelligence. Existing agents often struggle with robust long-term planning about …