paper-with-me

Papers

Cost-Matching Model Predictive Control for Efficient Reinforcement Learning in Humanoid Locomotion

2026-03-30 · Wenqi Cai, Kyriakos G. Vamvoudakis, Sébastien Gros, Anthony Tzes arxiv

In this paper, we propose a cost-matching approach for optimal humanoid locomotion within a Model Predictive Control (MPC)-based Reinforcement Learning (RL) framework. A parameterized MPC formulation with centroidal dynamics is trained to approximate the action-value function obtained from high-fidelity closed-loop data. Specifically, the MPC cost-to-go is evaluated along recorded state-action trajectories, and the parameters are updated to minimize the discrepancy between MPC-predicted values and measured returns. This formulation enables efficient gradient-based learning while avoiding the computational burden of repeatedly solving the MPC problem during training. The proposed method is validated in simulation using a commercial humanoid platform. Results demonstrate improved locomotion performance and robustness to model mismatch and external disturbances compared with manually tuned baselines.

📄 PDF Abstract BibTeX arXiv:2603.28243

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Predictive Style Matching: Natural and Robust Humanoid Locomotion

2026-06-05 · Simeon Nedelchev, Ekaterina Chaikovskaia, Egor Davydenko, Eduard Zaliaev 외 arxiv

Reinforcement learning has become the prevailing approach to humanoid locomotion control: policies transfer reliably from simulation to hardware and recover gracefully from disturbances. Motion quality, however, still la…

Reinforcement Learning

RGB: RL Guided Whole-Body MPPI for Humanoid Control

2026-06-23 · Yunsoo Seo, Sol Choi, Euncheol Im, Myo Taeg Lim 외 arxiv

Humanoid robots require whole-body controllers that are both robust and precise in contact-rich environments. While deep reinforcement learning (RL) achieves robust stability, its behavior is tightly coupled to the train…

Reinforcement Learning

Mobile-TeleVision: Predictive Motion Priors for Humanoid Whole-Body Control

2024-12-10 · Chenhao Lu, Xuxin Cheng, Jialong Li, Shiqi Yang 외

Humanoid robots require both robust lower-body locomotion and precise upper-body manipulation. While recent Reinforcement Learning (RL) approaches provide whole-body loco-manipulation policies, they lack precise manipula…

motion retargetingReinforcement Learning (RL)

Learning from Massive Human Videos for Universal Humanoid Pose Control

2024-12-18 · Jiageng Mao, Siheng Zhao, Siqi Song, Tianheng Shi 외

Scalable learning of humanoid robots is crucial for their deployment in real-world applications. While traditional approaches primarily rely on reinforcement learning or teleoperation to achieve whole-body control, they …

Caption GenerationHumanoid Controlmotion retargeting

RAVEN: Reinforcement-Adaptive Visibility-Graph Planning for Robust Humanoid Navigation with Collision-Free MPC

2026-07-17 · Ruochen Hou, Shiqi Wang, Beom Jun Kim, Hanzhang Fang 외 arxiv

Humanoid navigation in dynamic environments requires long-horizon planning while respecting short-horizon dynamic and safety constraints. Classical visibility-graph planners combined with model predictive control (MPC) c…

Hierarchical Reinforcement Learning