paper-with-me

홈 › Papers

Benchmarking Potential Based Rewards for Learning Humanoid Locomotion

2023-07-19 · Se Hwan Jeon, Steve Heim, Charles Khazoom, Sangbae Kim

The main challenge in developing effective reinforcement learning (RL) pipelines is often the design and tuning the reward functions. Well-designed shaping reward can lead to significantly faster learning. Naively formulated rewards, however, can conflict with the desired behavior and result in overfitting or even erratic performance if not properly tuned. In theory, the broad class of potential based reward shaping (PBRS) can help guide the learning process without affecting the optimal policy. Although several studies have explored the use of potential based reward shaping to accelerate learning convergence, most have been limited to grid-worlds and low-dimensional systems, and RL in robotics has predominantly relied on standard forms of reward shaping. In this paper, we benchmark standard forms of shaping with PBRS for a humanoid robot. We find that in this high-dimensional system, PBRS has only marginal benefits in convergence speed. However, the PBRS reward terms are significantly more robust to scaling than typical reward shaping approaches, and thus easier to tune.

📄 PDF Abstract BibTeX arXiv:2307.10142

Code (1)

se-hwan/pbrs-humanoid 공식 구현 pytorch

Tasks

BenchmarkingReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds

2025-02-14 · Huayi Wang, ZiRui Wang, Junli Ren, Qingwei Ben 외

Traversing risky terrains with sparse footholds poses a significant challenge for humanoid robots, requiring precise foot placements and stable locomotion. Existing learning-based approaches often struggle on such comple…

Reinforcement Learning (RL)

FLAM: Foundation Model-Based Body Stabilization for Humanoid Locomotion and Manipulation

2025-03-28 · Xianqi Zhang, Hongliang Wei, Wenrui Wang, Xingtao Wang 외

Humanoid robots have attracted significant attention in recent years. Reinforcement Learning (RL) is one of the main ways to control the whole body of humanoid robots. RL enables agents to complete tasks by learning from…

Reinforcement Learning (RL)

Humanoid Whole-Body Locomotion on Narrow Terrain via Dynamic Balance and Reinforcement Learning

2025-02-24 · Weiji Xie, Chenjia Bai, Jiyuan Shi, Junkai Yang 외

Humans possess delicate dynamic balance mechanisms that enable them to maintain stability across diverse terrains and under extreme conditions. However, despite significant advances recently, existing locomotion algorith…

Reinforcement Learning (RL)

Moving Through Clutter: Scaling Data Collection and Benchmarking for 3D Scene-Aware Humanoid Locomotion via Virtual Reality

2026-03-06 · Beichen Wang, Yuanjie Lu, Linji Wang, Liuchuan Yu 외 arxiv

Recent advances in humanoid locomotion have enabled dynamic behaviors such as dancing, martial arts, and parkour, yet these capabilities are predominantly demonstrated in open, flat, and obstacle-free settings. In contra…

No More Marching: Learning Humanoid Locomotion for Short-Range SE(2) Targets

2025-08-16 · Pranay Dugar, Mohitvishnu S. Gadde, Jonah Siekmann, Yesh Godse 외 arxiv

Humanoids operating in real-world workspaces must frequently execute task-driven, short-range movements to SE(2) target poses. To be practical, these transitions must be fast, robust, and energy efficient. While learning…

Reinforcement Learning