paper-with-me

홈 › Papers

Learning to Play by Imitating Humans

2020-06-11 · Rostam Dinyari, Pierre Sermanet, Corey Lynch

Acquiring multiple skills has commonly involved collecting a large number of expert demonstrations per task or engineering custom reward functions. Recently it has been shown that it is possible to acquire a diverse set of skills by self-supervising control on top of human teleoperated play data. Play is rich in state space coverage and a policy trained on this data can generalize to specific tasks at test time outperforming policies trained on individual expert task demonstrations. In this work, we explore the question of whether robots can learn to play to autonomously generate play data that can ultimately enhance performance. By training a behavioral cloning policy on a relatively small quantity of human play, we autonomously generate a large quantity of cloned play data that can be used as additional training. We demonstrate that a general purpose goal-conditioned policy trained on this augmented dataset substantially outperforms one trained only with the original human data on 18 difficult user-specified manipulation tasks in a simulated robotic tabletop environment. A video example of a robot imitating human play can be seen here: https://learning-to-play.github.io/videos/undirected_play1.mp4

📄 PDF Abstract BibTeX arXiv:2006.06874

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Transcendence: Generative Models Can Outperform The Experts That Train Them

2024-06-17 · Edwin Zhang, Vincent Zhu, Naomi Saphra, Anat Kleiman 외

Generative models are trained with the simple objective of imitating the conditional probability distribution induced by the data they are trained on. Therefore, when trained on data generated by humans, we may not expec…

Imitating careful experts to avoid catastrophic events

2023-02-02 · Jack R. P. Hanslope, Laurence Aitchison

RL is increasingly being used to control robotic systems that interact closely with humans. This interaction raises the problem of safe RL: how to ensure that a RL-controlled robotic system never, for instance, injures a…

UniSkill: Imitating Human Videos via Cross-Embodiment Skill Representations

2025-05-13 · Hanjung Kim, Jaehyun Kang, Hyolim Kang, Meedeum Cho 외

Mimicry is a fundamental learning mechanism in humans, enabling individuals to learn new tasks by observing and imitating experts. However, applying this ability to robots presents significant challenges due to the inher…

Atari-HEAD: Atari Human Eye-Tracking and Demonstration Dataset

2019-03-15 · Ruohan Zhang, Calen Walshe, Zhuode Liu, Lin Guan 외

Large-scale public datasets have been shown to benefit research in multiple areas of modern artificial intelligence. For decision-making research that requires human data, high-quality datasets serve as important benchma…

Decision MakingImitation LearningReinforcement Learning

Do humans and machines have the same eyes? Human-machine perceptual differences on image classification

2023-04-18 · Minghao Liu, Jiaheng Wei, Yang Liu, James Davis

Trained computer vision models are assumed to solve vision tasks by imitating human behavior learned from training labels. Most efforts in recent vision research focus on measuring the model task performance using standa…

image-classificationImage Classification