paper-with-me

홈 › Papers

Visual Imitation Made Easy

2020-08-11 · Sarah Young, Dhiraj Gandhi, Shubham Tulsiani, Abhinav Gupta, Pieter Abbeel, Lerrel Pinto

Visual imitation learning provides a framework for learning complex manipulation behaviors by leveraging human demonstrations. However, current interfaces for imitation such as kinesthetic teaching or teleoperation prohibitively restrict our ability to efficiently collect large-scale data in the wild. Obtaining such diverse demonstration data is paramount for the generalization of learned skills to novel scenarios. In this work, we present an alternate interface for imitation that simplifies the data collection process while allowing for easy transfer to robots. We use commercially available reacher-grabber assistive tools both as a data collection device and as the robot's end-effector. To extract action information from these visual demonstrations, we use off-the-shelf Structure from Motion (SfM) techniques in addition to training a finger detection network. We experimentally evaluate on two challenging tasks: non-prehensile pushing and prehensile stacking, with 1000 diverse demonstrations for each task. For both tasks, we use standard behavior cloning to learn executable policies from the previously collected offline demonstrations. To improve learning performance, we employ a variety of data augmentations and provide an extensive analysis of its effects. Finally, we demonstrate the utility of our interface by evaluating on real robotic scenarios with previously unseen objects and achieve a 87% success rate on pushing and a 62% success rate on stacking. Robot videos are available at https://dhiraj100892.github.io/Visual-Imitation-Made-Easy.

📄 PDF Abstract BibTeX arXiv:2008.04899

Code (0)

등록된 구현이 없습니다.

Tasks

Imitation Learning

Similar Papers 제목 키워드 기반

Dexterous Imitation Made Easy: A Learning-Based Framework for Efficient Dexterous Manipulation

2022-03-24 · Sridhar Pandian Arunachalam, Sneha Silwal, Ben Evans, Lerrel Pinto

Optimizing behaviors for dexterous manipulation has been a longstanding challenge in robotics, with a variety of methods from model-based control to model-free reinforcement learning having been previously explored in li…

Imitation Learning

Enhanced Motion-Text Alignment for Image-to-Video Transfer Learning

2024-01-01 · CVPR 2024 1 · Wei zhang, Chaoqun Wan, Tongliang Liu, Xinmei Tian 외

Extending large image-text pre-trained models (e.g. CLIP) for video understanding has made significant advancements. To enable the capability of CLIP to perceive dynamic information in videos existing works are dedic…

Transfer LearningVideo Understanding

Can VLMs Play Action Role-Playing Games? Take Black Myth Wukong as a Study Case

2024-09-19 · Peng Chen, Pi Bu, Jun Song, Yuan Gao 외

Recently, large language model (LLM)-based agents have made significant advances across various fields. One of the most popular research areas involves applying these agents to video games. Traditionally, these methods h…

Large Language Model

Cross-modal Multi-task Learning for Graphic Recognition of Caricature Face

2020-03-10 · Zuheng Ming, Jean-Christophe Burie, Muhammad Muzzamil Luqman

Face recognition of realistic visual images has been well studied and made a significant progress in the recent decade. Unlike the realistic visual images, the face recognition of the caricatures is far from the performa…

CaricatureFace RecognitionMulti-Task Learning

Imitation Learning from Pixel Observations for Continuous Control

2021-09-29 · samuel cohen, Brandon Amos, Marc Peter Deisenroth, Mikael Henaff 외

We study imitation learning using only visual observations for controlling dynamical systems with continuous states and actions. This setting is attractive due to the large amount of video data available from which agent…

Benchmarkingcontinuous-controlContinuous ControlImitation Learning