paper-with-me

홈 › Papers

Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids

2025-08-17 · Kaizhe Hu, Haochen Shi, Yao He, Weizhuo Wang, C. Karen Liu, Shuran Song arxiv

Simulation-based reinforcement learning (RL) has significantly advanced humanoid locomotion tasks, yet direct real-world RL from scratch or adapting from pretrained policies remains rare, limiting the full potential of humanoid robots. Real-world learning, despite being crucial for overcoming the sim-to-real gap, faces substantial challenges related to safety, reward design, and learning efficiency. To address these limitations, we propose Robot-Trains-Robot (RTR), a novel framework where a robotic arm teacher actively supports and guides a humanoid robot student. The RTR system provides protection, learning schedule, reward, perturbation, failure detection, and automatic resets. It enables efficient long-term real-world humanoid training with minimal human intervention. Furthermore, we propose a novel RL pipeline that facilitates and stabilizes sim-to-real transfer by optimizing a single dynamics-encoded latent variable in the real world. We validate our method through two challenging real-world humanoid tasks: fine-tuning a walking policy for precise speed tracking and learning a humanoid swing-up task from scratch, illustrating the promising capabilities of real-world humanoid learning realized by RTR-style systems. See https://robot-trains-robot.github.io/ for more info.

📄 PDF Abstract BibTeX arXiv:2508.12252

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

DayDreamer: World Models for Physical Robot Learning

2022-06-28 · Philipp Wu, Alejandro Escontrela, Danijar Hafner, Ken Goldberg 외

To solve tasks in complex environments, robots need to learn from experience. Deep reinforcement learning is a common approach to robot learning but requires a large amount of trial and error to learn, limiting its deplo…

Deep Reinforcement LearningNavigatereinforcement-learningReinforcement Learning (RL)

Robots that Collaborate: Sequential Asymmetric Imitation for Learning Coupled Robot Policies

2026-06-15 · Yincong Chen, Ranpeng Qiu, Zihao Li, Yanan Zhou 외 arxiv

Collaborative mobile manipulation requires robots to coordinate with a partially observed partner while physically interacting through shared objects. This is difficult because failures often arise not from poor local sk…

Robot Manipulation

Multi-Task Domain Adaptation for Deep Learning of Instance Grasping from Simulation

2017-10-17 · Kuan Fang, Yunfei Bai, Stefan Hinterstoisser, Silvio Savarese 외

Learning-based approaches to robotic manipulation are limited by the scalability of data collection and accessibility of labels. In this paper, we present a multi-task domain adaptation framework for instance grasping in…

Domain AdaptationInstance SegmentationSemantic SegmentationTransfer Learning

Strengthening Generative Robot Policies through Predictive World Modeling

2025-02-02 · Han Qi, Haocheng Yin, Aris Zhu, Yilun Du 외

We present generative predictive control (GPC), a learning control framework that (i) clones a generative diffusion-based policy from expert demonstrations, (ii) trains a predictive action-conditioned world model from bo…

UniDomain: Pretraining a Unified PDDL Domain from Real-World Demonstrations for Generalizable Robot Task Planning

2025-07-29 · Haoming Ye, Yunxiao Xiao, Cewu Lu, Panpan Cai arxiv

Robotic task planning in real-world environments requires reasoning over implicit constraints from language and vision. While LLMs and VLMs offer strong priors, they struggle with long-horizon structure and symbolic grou…

Robot Task PlanningRobot Manipulation