paper-with-me

Papers

PhysMoDPO: Physically-Plausible Humanoid Motion with Preference Optimization

2026-03-13 · Yangsong Zhang, Anujith Muraleedharan, Rikhat Akizhanov, Abdul Ahad Butt, Gül Varol, Pascal Fua, Fabio Pizzati, Ivan Laptev arxiv

Recent progress in text-conditioned human motion generation has been largely driven by diffusion models trained on large-scale human motion data. Building on this progress, recent methods attempt to transfer such models for character animation and real robot control by applying a Whole-Body Controller (WBC) that converts diffusion-generated motions into executable trajectories. While WBC trajectories become compliant with physics, they may expose substantial deviations from original motion. To address this issue, we here propose PhysMoDPO, a Direct Preference Optimization framework. Unlike prior work that relies on hand-crafted physics-aware heuristics such as foot-sliding penalties, we integrate WBC into our training pipeline and optimize diffusion model such that the output of WBC becomes compliant both with physics and original text instructions. To train PhysMoDPO we deploy physics-based and task-specific rewards and use them to assign preference to synthesized trajectories. Our extensive experiments on text-to-motion and spatial control tasks demonstrate consistent improvements of PhysMoDPO in both physical realism and task-related metrics on simulated robots. Moreover, we demonstrate that PhysMoDPO results in significant improvements when applied to zero-shot motion transfer in simulation and for real-world deployment on a G1 humanoid robot.

📄 PDF Abstract BibTeX arXiv:2603.13228

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SimGenHOI: Physically Realistic Whole-Body Humanoid-Object Interaction via Generative Modeling and Reinforcement Learning

2025-08-18 · Yuhang Lin, Yijia Xie, Jiahong Xie, Yuehao Huang 외 arxiv

Generating physically realistic humanoid-object interactions (HOI) is a fundamental challenge in robotics. Existing HOI generation approaches, such as diffusion-based models, often suffer from artifacts such as implausib…

Reinforcement Learning

It Takes Two: Learning Interactive Whole-Body Control Between Humanoid Robots

2025-10-11 · Zuhong Liu, Junhao Ge, Minhao Xiong, Jiahao Gu 외 arxiv

The true promise of humanoid robotics lies beyond single-agent autonomy: two or more humanoids must engage in physically grounded, socially meaningful whole-body interactions that echo the richness of human social intera…

Ground Reaction Inertial Poser: Physics-based Human Motion Capture from Sparse IMUs and Insole Pressure Sensors

2026-03-17 · Ryosuke Hori, Jyun-Ting Song, Zhengyi Luo, Jinkun Cao 외 arxiv

We propose Ground Reaction Inertial Poser (GRIP), a method that reconstructs physically plausible human motion using four wearable devices. Unlike conventional IMU-only approaches, GRIP combines IMU signals with foot pre…

PhysHMR: Learning Humanoid Control Policies from Vision for Physically Plausible Human Motion Reconstruction

2025-10-02 · Qiao Feng, Yiming Huang, Yufu Wang, Jiatao Gu 외 arxiv

Reconstructing physically plausible human motion from monocular videos remains a challenging problem in computer vision and graphics. Existing methods primarily focus on kinematics-based pose estimation, often leading to…

Reinforcement LearningPose Estimation

Towards Immersive Human-X Interaction: A Real-Time Framework for Physically Plausible Motion Synthesis

2025-08-04 · Kaiyang Ji, Ye Shi, Zichen Jin, Kangyi Chen 외 arxiv

Real-time synthesis of physically plausible human interactions remains a critical challenge for immersive VR/AR systems and humanoid robotics. While existing methods demonstrate progress in kinematic motion generation, t…

Reinforcement LearningMotion Synthesis