paper-with-me

Papers

From Experts to a Generalist: Toward General Whole-Body Control for Humanoid Robots

2025-06-15 · Yuxuan Wang, Ming Yang, Weishuai Zeng, Yu Zhang, Xinrun Xu, Haobin Jiang, Ziluo Ding, Zongqing Lu

Achieving general agile whole-body control on humanoid robots remains a major challenge due to diverse motion demands and data conflicts. While existing frameworks excel in training single motion-specific policies, they struggle to generalize across highly varied behaviors due to conflicting control requirements and mismatched data distributions. In this work, we propose BumbleBee (BB), an expert-generalist learning framework that combines motion clustering and sim-to-real adaptation to overcome these challenges. BB first leverages an autoencoder-based clustering method to group behaviorally similar motions using motion features and motion descriptions. Expert policies are then trained within each cluster and refined with real-world data through iterative delta action modeling to bridge the sim-to-real gap. Finally, these experts are distilled into a unified generalist controller that preserves agility and robustness across all motion types. Experiments on two simulations and a real humanoid robot demonstrate that BB achieves state-of-the-art general whole-body control, setting a new benchmark for agile, robust, and generalizable humanoid performance in the real world.

📄 PDF Abstract BibTeX arXiv:2506.12779

Code (0)

등록된 구현이 없습니다.

Tasks

Clustering

Similar Papers 제목 키워드 기반

Embodiment-Aware Generalist Specialist Distillation for Unified Humanoid Whole-Body Control

2026-02-03 · Quanquan Peng, Yunfeng Lin, Yufei Xue, Jiangmiao Pang 외 arxiv

Humanoid Whole-Body Controllers trained with reinforcement learning (RL) have recently achieved remarkable performance, yet many target a single robot embodiment. Variations in dynamics, degrees of freedom (DoFs), and ki…

Reinforcement Learning

Extreme-RGMT: Continual Learning of Highly Dynamic Skills for Robust Generalist Humanoid Control

2026-07-22 · Yubiao Ma, Han Yu, Kai Guo, Changtai Lv 외 arxiv

Humans can progressively acquire highly dynamic motor skills while preserving reliable everyday motor abilities. In contrast, existing humanoid controllers face a trade-off between generalist and specialist capabilities:…

Continual Learning

X-Loco: Towards Generalist Humanoid Locomotion Control via Synergetic Policy Distillation

2026-03-04 · Dewei Wang, Xinmiao Wang, Chenyun Zhang, Jiyuan Shi 외 arxiv

While recent advances have demonstrated strong performance in individual humanoid skills such as upright locomotion, fall recovery and whole-body coordination, learning a single policy that masters all these skills remai…

HAF: Adapting Generalist VLAs to Humanoid Whole-Body Loco-manipulation via Hierarchical Action Flow and Spectral Latent RL

2026-08-17 · Langzhe Gu, Chengkai Hou, Meng Li, Xinhua Wang 외 arxiv

Humanoid robots hold great promise as general-purpose agents in human-centered environments, yet generalist vision-language-action (VLA) foundation models are not readily applicable to humanoid whole-body loco-manipulati…

Dimensionality ReductionReinforcement Learning

HANDOFF: Humanoid Agentic Task-Space Whole-Body Control via Distilled Complementary Teachers

2026-06-04 · Lizhi Yang, Junheng Li, Nehar Poddar, Yiling Hou 외 arxiv

For a humanoid robot to be deployed in the real world, the choice of command space (i.e., the interface between task planning and whole-body control) is crucial. Existing whole-body controllers typically demand dense kin…