paper-with-me

Papers

RobotDancing: Residual-Action Reinforcement Learning Enables Robust Long-Horizon Humanoid Motion Tracking

2025-09-25 · Zhenguo Sun, Yibo Peng, Yuan Meng, Xukun Li, Bo-Sheng Huang, Zhenshan Bing, Xinlong Wang, Alois Knoll arxiv

Long-horizon, high-dynamic motion tracking on humanoids remains brittle because absolute joint commands cannot compensate model-plant mismatch, leading to error accumulation. We propose RobotDancing, a simple, scalable framework that predicts residual joint targets to explicitly correct dynamics discrepancies. The pipeline is end-to-end--training, sim-to-sim validation, and zero-shot sim-to-real--and uses a single-stage reinforcement learning (RL) setup with a unified observation, reward, and hyperparameter configuration. We evaluate primarily on Unitree G1 with retargeted LAFAN1 dance sequences and validate transfer on H1/H1-2. RobotDancing can track multi-minute, high-energy behaviors (jumps, spins, cartwheels) and deploys zero-shot to hardware with high motion tracking quality.

📄 PDF Abstract BibTeX arXiv:2509.20717

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Adaptive control of a mechatronic system using constrained residual reinforcement learning

2021-10-06 · Tom Staessens, Tom Lefebvre, Guillaume Crevecoeur

We propose a simple, practical and intuitive approach to improve the performance of a conventional controller in uncertain environments using deep reinforcement learning while maintaining safe operation. Our approach is …

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

From Prior to Pro: Efficient Skill Mastery via Distribution Contractive RL Finetuning

2026-03-10 · Zhanyi Sun, Shuran Song arxiv

We introduce Distribution Contractive Reinforcement Learning (DICE-RL), a framework that uses reinforcement learning (RL) as a "distribution contraction" operator to refine pretrained generative robot policies. DICE-RL t…

Reinforcement LearningSkill Mastery

Emergent temporal abstractions in autoregressive models enable hierarchical reinforcement learning

2025-12-23 · Seijin Kobayashi, Yanick Schimpf, Maximilian Schlegel, Angelika Steger 외 arxiv

Large-scale autoregressive models pretrained on next-token prediction and finetuned with reinforcement learning (RL) have achieved unprecedented success on many problem domains. During RL, these models explore by generat…

Hierarchical Reinforcement Learning

What Makes Value Learning Efficient in Residual Reinforcement Learning?

2026-02-11 · Guozheng Ma, Lu Li, Haoyu Wang, Zixuan Liu 외 arxiv

Residual reinforcement learning (RL) enables stable online refinement of expressive pretrained policies by freezing the base and learning only bounded corrections. However, value learning in residual RL poses unique chal…

Reinforcement Learning

Efficient Residual Learning with Mixture-of-Experts for Universal Dexterous Grasping

2024-10-03 · Ziye Huang, Haoqi Yuan, Yuhui Fu, Zongqing Lu

Universal dexterous grasping across diverse objects presents a fundamental yet formidable challenge in robot learning. Existing approaches using reinforcement learning (RL) to develop policies on extensive object dataset…

GPUMixture-of-ExpertsMulti-Task LearningReinforcement Learning (RL)