paper-with-me

홈 › Papers

Out-of-Dynamics Imitation Learning from Multimodal Demonstrations

2022-11-13 · Yiwen Qiu, Jialong Wu, Zhangjie Cao, Mingsheng Long

Existing imitation learning works mainly assume that the demonstrator who collects demonstrations shares the same dynamics as the imitator. However, the assumption limits the usage of imitation learning, especially when collecting demonstrations for the imitator is difficult. In this paper, we study out-of-dynamics imitation learning (OOD-IL), which relaxes the assumption to that the demonstrator and the imitator have the same state spaces but could have different action spaces and dynamics. OOD-IL enables imitation learning to utilize demonstrations from a wide range of demonstrators but introduces a new challenge: some demonstrations cannot be achieved by the imitator due to the different dynamics. Prior works try to filter out such demonstrations by feasibility measurements, but ignore the fact that the demonstrations exhibit a multimodal distribution since the different demonstrators may take different policies in different dynamics. We develop a better transferability measurement to tackle this newly-emerged challenge. We firstly design a novel sequence-based contrastive clustering algorithm to cluster demonstrations from the same mode to avoid the mutual interference of demonstrations from different modes, and then learn the transferability of each demonstration with an adversarial-learning based algorithm in each cluster. Experiment results on several MuJoCo environments, a driving environment, and a simulated robot environment show that the proposed transferability measurement more accurately finds and down-weights non-transferable demonstrations and outperforms prior works on the final imitation learning performance. We show the videos of our experiment results on our website.

📄 PDF Abstract BibTeX arXiv:2211.06839

Code (1)

evieq01/oodil 공식 구현 pytorch

Tasks

Imitation LearningMuJoCo

Similar Papers 제목 키워드 기반

Learning from Imperfect Demonstrations from Agents with Varying Dynamics

2021-03-10 · Zhangjie Cao, Dorsa Sadigh

Imitation learning enables robots to learn from demonstrations. Previous imitation learning algorithms usually assume access to optimal expert demonstrations. However, in many real-world applications, this assumption is …

Imitation Learning

Zero-shot Imitation Learning from Demonstrations for Legged Robot Visual Navigation

2019-09-27 · Xinlei Pan, Tingnan Zhang, Brian Ichter, Aleksandra Faust 외

Imitation learning is a popular approach for training visual navigation policies. However, collecting expert demonstrations for legged robots is challenging as these robots can be hard to control, move slowly, and cannot…

DisentanglementImitation LearningVisual Navigation

Imitation Learning for Fashion Style Based on Hierarchical Multimodal Representation

2020-04-13 · Shizhu Liu, Shanglin Yang, Hui Zhou

Fashion is a complex social phenomenon. People follow fashion styles from demonstrations by experts or fashion icons. However, for machine agent, learning to imitate fashion experts from demonstrations can be challenging…

Imitation LearningReinforcement Learning

Domain Adaptive Imitation Learning

2019-09-30 · ICML 2020 1 · Kuno Kim, Yihong Gu, Jiaming Song, Shengjia Zhao 외

We study the question of how to imitate tasks across domains with discrepancies such as embodiment, viewpoint, and dynamics mismatch. Many prior works require paired, aligned demonstrations and an additional RL step that…

Imitation Learning

Offline Imitation Learning with a Misspecified Simulator

2020-12-01 · NeurIPS 2020 12 · Shengyi Jiang, JingCheng Pang, Yang Yu

In real-world decision-making tasks, learning an optimal policy without a trial-and-error process is an appealing challenge. When expert demonstrations are available, imitation learning that mimics expert actions can le…

Decision MakingFrictionImitation LearningMuJoCo