paper-with-me

Papers

BFM-Zero: A Promptable Behavioral Foundation Model for Humanoid Control Using Unsupervised Reinforcement Learning

2025-11-06 · Yitang Li, Zhengyi Luo, Tonghe Zhang, Cunxi Dai, Anssi Kanervisto, Andrea Tirinzoni, Haoyang Weng, Kris Kitani, Mateusz Guzek, Ahmed Touati, Alessandro Lazaric, Matteo Pirotta, Guanya Shi arxiv

Building Behavioral Foundation Models (BFMs) for humanoid robots has the potential to unify diverse control tasks under a single, promptable generalist policy. However, existing approaches are either exclusively deployed on simulated humanoid characters, or specialized to specific tasks such as tracking. We propose BFM-Zero, a framework that learns an effective shared latent representation that embeds motions, goals, and rewards into a common space, enabling a single policy to be prompted for multiple downstream tasks without retraining. This well-structured latent space in BFM-Zero enables versatile and robust whole-body skills on a Unitree G1 humanoid in the real world, via diverse inference methods, including zero-shot motion tracking, goal reaching, and reward optimization, and few-shot optimization-based adaptation. Unlike prior on-policy reinforcement learning (RL) frameworks, BFM-Zero builds upon recent advancements in unsupervised RL and Forward-Backward (FB) models, which offer an objective-centric, explainable, and smooth latent representation of whole-body motions. We further extend BFM-Zero with critical reward shaping, domain randomization, and history-dependent asymmetric learning to bridge the sim-to-real gap. Those key design choices are quantitatively ablated in simulation. A first-of-its-kind model, BFM-Zero establishes a step toward scalable, promptable behavioral foundation models for whole-body humanoid control.

📄 PDF Abstract BibTeX arXiv:2511.04131

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Behavior Foundation Model for Humanoid Robots

2025-09-17 · Weishuai Zeng, Shunlin Lu, Kangning Yin, Xiaojie Niu 외 arxiv

Whole-body control (WBC) of humanoid robots has witnessed remarkable progress in skill versatility, enabling a wide range of applications such as locomotion, teleoperation, and motion tracking. Despite these achievements…

Scaling Behavior Foundation Model for Humanoid Robots

2026-07-16 · Weishuai Zeng, Kangning Yin, Xiaojie Niu, Shunlin Lu 외 arxiv

Humanoid control requires natural whole-body coordination, precise real-time responses to control signals, and robust generalization across diverse environmental contexts, making it a cornerstone for generalist embodied …

Zero-Shot Whole-Body Humanoid Control via Behavioral Foundation Models

2025-04-15 · Andrea Tirinzoni, Ahmed Touati, Jesse Farebrother, Mateusz Guzek 외

Unsupervised reinforcement learning (RL) aims at pre-training agents that can solve a wide range of downstream tasks in complex environments. Despite recent advancements, existing approaches suffer from several limitatio…

Humanoid ControlReinforcement Learning (RL)Unsupervised Reinforcement LearningZero-shot Generalization

Learning to Get Up Across Morphologies: Zero-Shot Recovery with a Unified Humanoid Policy

2025-12-13 · Jonathan Spraggett arxiv

Fall recovery is a critical skill for humanoid robots in dynamic environments such as RoboCup, where prolonged downtime often decides the match. Recent techniques using deep reinforcement learning (DRL) have produced rob…

Zero-shot GeneralizationReinforcement Learning

HoloMotion-1 Technical Report

2026-05-14 · Maiyue Chen, Kaihui Wang, Bo Zhang, Xihan Ma 외 arxiv

In this report, we present HoloMotion-1, a humanoid motion foundation model for zero-shot whole-body motion tracking. A key innovation of HoloMotion-1 is to scale control-policy training with a large-scale hybrid motion …