paper-with-me

홈 › Papers

HOST:Robots Acquire Manipulation Skills in Seconds from a Single Human Video

2026-07-22 · Guangyan Chen, Meiling Wang, Te Cui, Zichen Zhou, Qi Shao, Shalfun Li, Hang Su, Roy Gan, Hao Wang, Mengyin Fu, Yi Yang, Yufeng Yue arxiv

The ability to acquire skills rapidly and effortlessly while retaining those already mastered is essential for robots. However, current methods still rely on a cumbersome training-time loop that is costly and slow, while eroding skills already mastered. In this paper, we introduce HOST (Human-to-robot One-Shot Skill AcquisiTion), a framework that enables a robot to acquire skills in seconds from a single human video while retaining previously mastered skills. HOST resolves skill acquisition through a cascade of self-grounded prediction. It first estimates the robot's progress within the demonstrated task, then translates the upcoming progression into the robot's own future observations, and finally derives actions from these predicted observations. This cascade is trained on targets coupled to the video demonstration, obtained by mapping the robot trajectory and the video demonstration onto a shared task progress manifold, then redefining each target to align with the future progression of the video. HOST thereby enables the robot to actively follow the demonstrated procedure and adapt it to the robot's embodiment. HOST acquires novel skills at inference time from a single human video in an average of 29 seconds and achieves a 62% average success rate. It exceeds the zero-shot baseline by 45% while retaining previously mastered skills. HOST even exceeds the baseline fine-tuned on 50 robot demonstrations per task while requiring 50 times fewer demonstrations and acquiring each skill 507 times faster. Additional information about HOST is available on the project website.

📄 PDF Abstract BibTeX arXiv:2607.20033

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Simple Approach to Continual Learning by Transferring Skill Parameters

2021-10-19 · K. R. Zentner, Ryan Julian, Ujjwal Puri, Yulun Zhang 외

In order to be effective general purpose machines in real world environments, robots not only will need to adapt their existing manipulation skills to new circumstances, they will need to acquire entirely new skills on-t…

Continual Learning

RH20T: A Comprehensive Robotic Dataset for Learning Diverse Skills in One-Shot

2023-07-02 · Hao-Shu Fang, Hongjie Fang, Zhenyu Tang, Jirong Liu 외

A key challenge in robotic manipulation in open domains is how to acquire diverse and generalizable skills for robots. Recent research in one-shot imitation learning has shown promise in transferring trained policies to …

Imitation LearningMotion PlanningRobot ManipulationTask and Motion Planning

Enhancing Reusability of Learned Skills for Robot Manipulation via Gaze and Bottleneck

2025-02-25 · Ryo Takizawa, Izumi Karino, Koki Nakagawa, Yoshiyuki Ohmura 외

Autonomous agents capable of diverse object manipulations should be able to acquire a wide range of manipulation skills with high reusability. Although advances in deep learning have made it increasingly feasible to repl…

Imitation LearningObjectRobot Manipulation

Generalizable Humanoid Manipulation with 3D Diffusion Policies

2024-10-14 · Yanjie Ze, Zixuan Chen, Wenhao Wang, Tianyi Chen 외

Humanoid robots capable of autonomous operation in diverse environments have long been a goal for roboticists. However, autonomous manipulation by humanoid robots has largely been restricted to one specific scene, primar…

Camera CalibrationPoint Cloud Segmentation

WildLMa: Long Horizon Loco-Manipulation in the Wild

2024-11-22 · Ri-Zhao Qiu, Yuchen Song, Xuanbin Peng, Sai Aneesh Suryadevara 외

'In-the-wild' mobile manipulation aims to deploy robots in diverse real-world environments, which requires the robot to (1) have skills that generalize across object configurations; (2) be capable of long-horizon task ex…

Imitation Learning