paper-with-me

Papers

Learning a Multi-Modal Policy via Imitating Demonstrations with Mixed Behaviors

2019-03-25 · Fang-I Hsiao, Jui-Hsuan Kuo, Min Sun

We propose a novel approach to train a multi-modal policy from mixed demonstrations without their behavior labels. We develop a method to discover the latent factors of variation in the demonstrations. Specifically, our method is based on the variational autoencoder with a categorical latent variable. The encoder infers discrete latent factors corresponding to different behaviors from demonstrations. The decoder, as a policy, performs the behaviors accordingly. Once learned, the policy is able to reproduce a specific behavior by simply conditioning on a categorical vector. We evaluate our method on three different tasks, including a challenging task with high-dimensional visual inputs. Experimental results show that our approach is better than various baseline methods and competitive with a multi-modal policy trained by ground truth behavior labels.

📄 PDF Abstract BibTeX arXiv:1903.10304

Code (0)

등록된 구현이 없습니다.

Tasks

Decoder

Methods 이 논문이 사용한 방법론

Solana Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

Trajectory VAE for multi-modal imitation

2019-05-01 · ICLR 2019 5 · Xiaoyu Lu, Jan Stuehmer, Katja Hofmann

We address the problem of imitating multi-modal expert demonstrations in sequential decision making problems. In many practical applications, for example video games, behavioural demonstrations are readily available that…

continuous-controlContinuous ControlDecision MakingImitation Learning+2

Imitating Human Behaviour with Diffusion Models

2023-01-25 · Tim Pearce, Tabish Rashid, Anssi Kanervisto, Dave Bignell 외

Diffusion models have emerged as powerful generative models in the text-to-image domain. This paper studies their application as observation-to-action models for imitating human behaviour in sequential environments. Huma…

GROOT-2: Weakly Supervised Multi-Modal Instruction Following Agents

2024-12-07 · Shaofei Cai, Bowei Zhang, ZiHao Wang, Haowei Lin 외

Developing agents that can follow multimodal instructions remains a fundamental challenge in robotics and AI. Although large-scale pre-training on unlabeled datasets (no language instruction) has enabled agents to learn …

Instruction Following

Learning to Play by Imitating Humans

2020-06-11 · Rostam Dinyari, Pierre Sermanet, Corey Lynch

Acquiring multiple skills has commonly involved collecting a large number of expert demonstrations per task or engineering custom reward functions. Recently it has been shown that it is possible to acquire a diverse set …

Self-Imitation Learning via Trajectory-Conditioned Policy for Hard-Exploration Tasks

2019-09-25 · Yijie Guo, Jongwook Choi, Marcin Moczulski, Samy Bengio 외

Imitation learning from human-expert demonstrations has been shown to be greatly helpful for challenging reinforcement learning problems with sparse environment rewards. However, it is very difficult to achieve similar s…

Imitation Learning