paper-with-me

Papers

OpenHLM: An Empirical Recipe for Whole-Body Humanoid Loco-Manipulation

2026-06-20 · Yingdong Hu, Haodong Zhu, Boyuan Zheng, Yihang Hu, Tong Zhang, Zunhao Chen, Junming Zhao, Ruiqian Nai, Yang Gao arxiv

Whole-body humanoid loco-manipulation requires coordinating the robot's entire kinematic chain. However, most existing systems typically decouple the upper and lower bodies into separate controllers, limiting such coordination and yielding behaviors similar to those of a wheeled dual-arm platform. In this paper, we ask what it takes to build a whole-body native vision-language-action (VLA) model that maps language and pixels directly to all of the humanoid's degrees of freedom. We conduct a systematic empirical study organized as a roadmap of one-variable-at-a-time experiments across three phases: whole-body teleoperation, VLA model design, and heterogeneous co-training. Our study yields several intriguing findings: a joint-based whole-body teleoperation interface outperforms alternatives that only partially expose the humanoid's degrees of freedom; a VLA pretrained on static and wheeled dual-arm platforms transfers surprisingly well to a humanoid's full action space; and co-training with HuMI, the humanoid analog of UMI, extends the policy to new objects and instructions without additional whole-body teleoperation on those targets. Following this roadmap yields OpenHLM, an open-source recipe for whole-body humanoid loco-manipulation. In a challenging long-horizon task that spans a wide vertical range of the humanoid, OpenHLM outperforms two state-of-the-art humanoid VLA baselines (GR00T N1.6 and $Ψ_0$) using less than half the total demonstration time. Our code, training data, and model checkpoints are available at [https://openhlm-project.github.io/].

📄 PDF Abstract BibTeX arXiv:2606.22174

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Athena-WBC: Capability-Aligned Policy Experts for Long-Tail Humanoid Whole-Body Control

2026-07-06 · Yuan Jiang, Ningyuan Zhang, Xicun Yang, Shidi Li 외 arxiv

Large-scale humanoid motion-tracking controllers are commonly improved by reallocating training effort: difficult motions are sampled more often, isolated into smaller subsets, or assigned to specialized experts. We show…

HumanoidExo: Scalable Whole-Body Humanoid Manipulation via Wearable Exoskeleton

2025-10-03 · Rui Zhong, Yizhe Sun, Junjie Wen, Jinming Li 외 arxiv

A significant bottleneck in humanoid policy learning is the acquisition of large-scale, diverse datasets, as collecting reliable real-world data remains both difficult and cost-prohibitive. To address this limitation, we…

Learning Sim-to-Real Humanoid Locomotion in 15 Minutes

2025-12-01 · Younggyo Seo, Carmelo Sferrazza, Juyue Chen, Guanya Shi 외 arxiv

Massively parallel simulation has reduced reinforcement learning (RL) training time for robots from days to minutes. However, achieving fast and reliable sim-to-real RL for humanoid control remains difficult due to the c…

Reinforcement Learning

OmniClone: Engineering a Robust, All-Rounder Whole-Body Humanoid Teleoperation System

2026-03-15 · Yixuan Li, Le Ma, Yutang Lin, Yushi Du 외 arxiv

Whole-body humanoid teleoperation enables humans to remotely control humanoid robots, serving as both a real-time operational tool and a scalable engine for collecting demonstrations for autonomous learning. Despite rece…

ZeroWBC: Learning Natural Whole-Body Humanoid Interaction from Human Egocentric Data

2026-03-10 · Haoran Yang, Jiacheng Bao, Yucheng Xin, Haoming Song 외 arxiv

Achieving versatile and natural whole-body humanoid interaction control remains challenging due to the high cost of whole-body teleoperation data. We present ZeroWBC, a teleoperation-free framework that learns humanoid w…