paper-with-me

홈 › Papers

Learning Multi-Modal Whole-Body Control for Real-World Humanoid Robots

2024-07-30 · Pranay Dugar, Aayam Shrestha, Fangzhou Yu, Bart van Marum, Alan Fern

The foundational capabilities of humanoid robots should include robustly standing, walking, and mimicry of whole and partial-body motions. This work introduces the Masked Humanoid Controller (MHC), which supports all of these capabilities by tracking target trajectories over selected subsets of humanoid state variables while ensuring balance and robustness against disturbances. The MHC is trained in simulation using a carefully designed curriculum that imitates partially masked motions from a library of behaviors spanning standing, walking, optimized reference trajectories, re-targeted video clips, and human motion capture data. It also allows for combining joystick-based control with partial-body motion mimicry. We showcase simulation experiments validating the MHC's ability to execute a wide variety of behaviors from partially-specified target motions. Moreover, we demonstrate sim-to-real transfer on the real-world Digit V3 humanoid robot. To our knowledge, this is the first instance of a learned controller that can realize whole-body control of a real-world humanoid for such diverse multi-modal targets.

📄 PDF Abstract BibTeX arXiv:2408.07295

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Library 설명 없음

Similar Papers 제목 키워드 기반

M3imic: Learning a Versatile Whole-Body Controller for Multimodal Motion Mimicking

2026-06-03 · Zuxing Lu, Ziang Zheng, Yao Lyu, Jingyu Liu 외 arxiv

Building a general-purpose whole-body controller is essential for enabling diverse motion capabilities in humanoid robots across a wide range of downstream tasks, including locomotion and loco-manipulation. Different tas…

Reinforcement Learning

MotionCraft: Crafting Whole-Body Motion with Plug-and-Play Multimodal Controls

2024-07-30 · Yuxuan Bian, Ailing Zeng, Xuan Ju, Xian Liu 외

Whole-body multimodal motion generation, controlled by text, speech, or music, has numerous applications including video generation and character animation. However, employing a unified model to achieve various generatio…

Gesture GenerationMotion GenerationMotion Synthesismultimodal generation

Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos

2025-08-12 · Chaoyi Wang, Yifan Yang, Jun Pei, Lijie Xia 외 arxiv

Creating realistic, fully animatable whole-body avatars from a single portrait is challenging due to limitations in capturing subtle expressions, body movements, and dynamic backgrounds. Current evaluation datasets and m…

SENTINEL: A Fully End-to-End Language-Action Model for Humanoid Whole Body Control

2025-11-24 · Yuxuan Wang, Haobin Jiang, Shiqing Yao, Ziluo Ding 외 arxiv

Existing humanoid control systems often rely on teleoperation or modular generation pipelines that separate language understanding from physical execution. However, the former is entirely human-driven, and the latter lac…

Versatile Multimodal Controls for Expressive Talking Human Animation

2025-03-10 · Zheng Qin, Ruobing Zheng, Yabing Wang, Tianqi Li 외

In filmmaking, directors typically allow actors to perform freely based on the script before providing specific guidance on how to present key actions. AI-generated content faces similar requirements, where users not onl…

Human Animation