paper-with-me

홈 › Papers

Morphology-Consistent Humanoid Interaction through Robot-Centric Video Synthesis

2026-03-20 · Weisheng Xu, Jian Li, Yi Gu, Bin Yang, Haodong Chen, Shuyi Lin, Mingqian Zhou, Jing Tan, Qiwei Wu, Xiangrui Jiang, Taowen Wang, Jiawen Wen, Qiwei Liang, Jiaxi Zhang, Renjing Xu arxiv

Equipping humanoid robots with versatile interaction skills typically requires either extensive policy training or explicit human-to-robot motion retargeting. However, learning-based policies face prohibitive data collection costs. Meanwhile, retargeting relies on human-centric pose estimation (e.g., SMPL), introducing a morphology gap. Skeletal scale mismatches result in severe spatial misalignments when mapped to robots, compromising interaction success. In this work, we propose Dream2Act, a robot-centric framework enabling zero-shot interaction through generative video synthesis. Given a third-person image of the robot and target object, our framework leverages video generation models to envision the robot completing the task with morphology-consistent motion. We employ a high-fidelity pose extraction system to recover physically feasible, robot-native joint trajectories from these synthesized dreams, subsequently executed via a general-purpose whole-body controller. Operating strictly within the robot-native coordinate space, Dream2Act avoids retargeting errors and eliminates task-specific policy training. We evaluate Dream2Act on the Unitree G1 across four whole-body mobile interaction tasks: ball kicking, sofa sitting, bag punching, and box hugging. Dream2Act achieves a 37.5% overall success rate, compared to 0% for conventional retargeting. While retargeting fails to establish correct physical contacts due to the morphology gap (with errors compounded during locomotion), Dream2Act maintains robot-consistent spatial alignment, enabling reliable contact formation and substantially higher task completion.

📄 PDF Abstract BibTeX arXiv:2603.19709

Code (0)

등록된 구현이 없습니다.

Tasks

Video GenerationPose Estimation

Similar Papers 제목 키워드 기반

Toward Humanoid Brain-Body Co-design: Joint Optimization of Control and Morphology for Fall Recovery

2025-10-25 · Bo Yue, Sheng Xu, Kui Jia, Guiliang Liu arxiv

Humanoid robots represent a central frontier in embodied intelligence, as their anthropomorphic form enables natural deployment in humans' workspace. Brain-body co-design for humanoids presents a promising approach to re…

Human2Humanoid: Physics-Aware Cross-Morphology Motion Retargeting for Humanoid Robots

2026-06-02 · Tianchen Huang, Feiyang Yuan, Junchi Gu, Shurui Fang 외 arxiv

Retargeting human motion to humanoid robots is critical for teleoperation, imitation learning and human-robot interaction. However, it remains challenging because of substantial morphological discrepancies between humans…

Learning Whole-Body Human-Humanoid Interaction from Human-Human Demonstrations

2026-01-14 · Wei-Jin Huang, Yue-Yi Zhang, Yi-Lin Wei, Zhi-Wei Xia 외 arxiv

Enabling humanoid robots to physically interact with humans is a critical frontier, but progress is hindered by the scarcity of high-quality Human-Humanoid Interaction (HHoI) data. While leveraging abundant Human-Human I…

Towards Shared Embodied Intelligence in Humanoid Robots through Optimization Development and Testing of the Human Aware ergoCub Robot

2026-05-26 · Carlotta Sartore, Mohamed Elobaid, Lorenzo Rapetti, Giulio Romualdi 외 arxiv

Collaboration is central to human behavior, enabling tasks beyond individual capability. This ability arises from coordinating actions through internal representations of others, a concept known as shared intelligence. A…

Handroid: Bridging Dexterous Hand and Humanoid

2026-07-17 · Ruogu Li, Chenyang Ma, Sikai Li, Zhenyu Wei 외 arxiv

Dexterous hands and humanoid robots are typically developed as distinct embodiments: the former enable contact-rich manipulation at the object scale, whereas the latter provide mobility and whole-body interaction in huma…