paper-with-me

홈 › Papers

RoboCLIP: One Demonstration is Enough to Learn Robot Policies

2023-10-11 · NeurIPS 2023 11

Reward specification is a notoriously difficult problem in reinforcement learning, requiring extensive expert supervision to design robust reward functions. Imitation learning (IL) methods attempt to circumvent these problems by utilizing expert demonstrations but typically require a large number of in-domain expert demonstrations. Inspired by advances in the field of Video-and-Language Models (VLMs), we present RoboCLIP, an online imitation learning method that uses a single demonstration (overcoming the large data requirement) in the form of a video demonstration or a textual description of the task to generate rewards without manual reward function design. Additionally, RoboCLIP can also utilize out-of-domain demonstrations, like videos of humans solving the task for reward generation, circumventing the need to have the same demonstration and deployment domains. RoboCLIP utilizes pretrained VLMs without any finetuning for reward generation. Reinforcement learning agents trained with RoboCLIP rewards demonstrate 2-3 times higher zero-shot performance than competing imitation learning methods on downstream robot manipulation tasks, doing so using only one video/text demonstration.

📄 PDF Abstract BibTeX arXiv:2310.07899

Code (1)

sumedh7/RoboCLIP pytorch

Tasks

Imitation Learningreinforcement-learningReinforcement LearningRobot Manipulation

Similar Papers 제목 키워드 기반

RoboTok: An Internet-Scale Data Engine for Human Demonstration Retrieval and Dexterous Manipulation Learning

2026-09-02 · Howard Qian, Yiting Chen, Yunfei Xie, Kejia Ren 외 hf

Robot learning increasingly depends on broad and diverse demonstrations, yet collecting robot data remains expensive and poorly suited to covering the long tail of real-world tasks. To address this bottleneck, we introdu…

Learning Modular Robot Locomotion from Demonstrations

2022-10-31 · Julian Whitman, Howie Choset

Modular robots can be reconfigured to create a variety of designs from a small set of components. But constructing a robot's hardware on its own is not enough -- each robot needs a controller. One could create controller…

Graph Neural NetworkImitation Learning

Deep Reinforcement Learning for Robotic Manipulation with Asynchronous Off-Policy Updates

2016-10-03 · Shixiang Gu, Ethan Holly, Timothy Lillicrap, Sergey Levine

Reinforcement learning holds the promise of enabling autonomous robots to learn large repertoires of behavioral skills with minimal human intervention. However, robotic applications of reinforcement learning often compro…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

One Demonstration Is Enough for Real-World Robotic Reinforcement Learning

2026-07-02 · Yuwan Liu, Hongze Yu, Song Liu, Yuhan Wang 외 arxiv

Learning effective robot control policies on physical hardware is challenging due to costly data collection and the difficulty of reward specification. Prior work has incorporated demonstrations into reinforcement learni…

Reinforcement Learning

MonoDuo: Using One Robot Arm to Learn Bimanual Policies

2026-05-28 · Sandeep Bajamahal, Lawrence Yunliang Chen, Toru Lin, Zehan Ma 외 arxiv

Bimanual coordination is essential for many real-world manipulation tasks, yet learning bimanual robot policies is limited by the scarcity of bimanual robots and datasets. Single-arm robots, however, are widely available…

Point Cloud SegmentationHand Pose Estimation