paper-with-me

홈 › Papers

VisualMimic: Visual Humanoid Loco-Manipulation via Motion Tracking and Generation

2025-09-24 · Shaofeng Yin, Yanjie Ze, Hong-Xing Yu, C. Karen Liu, Jiajun Wu arxiv

Humanoid loco-manipulation in unstructured environments demands tight integration of egocentric perception and whole-body control. However, existing approaches either depend on external motion capture systems or fail to generalize across diverse tasks. We introduce VisualMimic, a visual sim-to-real framework that unifies egocentric vision with hierarchical whole-body control for humanoid robots. VisualMimic combines a task-agnostic low-level keypoint tracker -- trained from human motion data via a teacher-student scheme -- with a task-specific high-level policy that generates keypoint commands from visual and proprioceptive input. To ensure stable training, we inject noise into the low-level policy and clip high-level actions using human motion statistics. VisualMimic enables zero-shot transfer of visuomotor policies trained in simulation to real humanoid robots, accomplishing a wide range of loco-manipulation tasks such as box lifting, pushing, football dribbling, and kicking. Beyond controlled laboratory settings, our policies also generalize robustly to outdoor environments. Videos are available at: https://visualmimic.github.io .

📄 PDF Abstract BibTeX arXiv:2509.20322

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MotionWAM: Towards Foundation World Action Models for Real-Time Humanoid Loco-Manipulation

2026-06-08 · Jia Zheng, Teli Ma, Yudong Fan, Zifan Wang 외 arxiv

World Action Models (WAMs) couple a video dynamics prior to the policy and have shown encouraging results on tabletop manipulation, but iterative denoising over high-dimensional video-action latents leaves them too slow …

$ω$-0: A Latent Predictive World Action Model for Concurrent Humanoid Loco-Manipulation

2026-08-06 · Zhe Li, Zhenzhe Zhang, Yangyang Wei, Wenjie Zhang 외 arxiv

Humanoid household tasks often require concurrent loco-manipulation, where the robot must move, adjust posture, maintain balance, and manipulate objects as a single coordinated behavior. Yet existing humanoid policies ty…

WholeBodyVLA: Towards Unified Latent VLA for Whole-Body Loco-Manipulation Control

2025-12-11 · Haoran Jiang, Jin Chen, Qingwen Bu, Li Chen 외 arxiv

Humanoid robots require precise locomotion and dexterous manipulation to perform challenging loco-manipulation tasks. Yet existing approaches, modular or end-to-end, are deficient in manipulation-aware locomotion. This c…

Mobile-TeleVision: Predictive Motion Priors for Humanoid Whole-Body Control

2024-12-10 · Chenhao Lu, Xuxin Cheng, Jialong Li, Shiqi Yang 외

Humanoid robots require both robust lower-body locomotion and precise upper-body manipulation. While recent Reinforcement Learning (RL) approaches provide whole-body loco-manipulation policies, they lack precise manipula…

motion retargetingReinforcement Learning (RL)

DemoHLM: From One Demonstration to Generalizable Humanoid Loco-Manipulation

2025-10-13 · Yuhui Fu, Feiyang Xie, Chaoyi Xu, Jing Xiong 외 arxiv

Loco-manipulation is a fundamental challenge for humanoid robots to achieve versatile interactions in human environments. Although recent studies have made significant progress in humanoid whole-body control, loco-manipu…