paper-with-me

홈 › Papers

Progress Reward Modeling for Robotic Learning: A Comprehensive Survey

2026-07-22 · Jianshu Zhang, Keliang Wu, Haoran Lu, Anbang Liu, Ce Zhang, Weijie Yin, Chengxuan Qian, Xiyuan Yang, Zhenyu Pan, Guo Ye, Han Liu hf

Robotic learning takes place in dynamic environments with large behavior spaces. A terminal success signal only tells the robot whether the task is completed. It does not explain whether the current behavior is making progress, remaining unchanged, or undoing earlier progress. For this reason, recent studies have increasingly explored progress rewards that provide feedback during task execution. However, the current literature lacks a shared framework. Existing methods use different observations, goal specifications, output signals, supervision sources, and evaluation protocols. This makes it difficult to compare them and understand what their results actually validate. In this survey, we provide a unified view of progress reward modeling for robotic learning. We organize the field in three connected steps. We first study the interface of a progress model. This defines the problem from the outside by asking what information the model receives and what form of progress signal it produces. We then move inside the model and study the methods used to construct this signal. This reveals the different assumptions and mechanisms behind progress estimation and reward generation. Finally, we examine the data and benchmarks that support these methods. This shows how progress supervision is obtained and what different evaluations actually measure. Together, these three perspectives connect what a progress model is, how it is built, and how its quality is validated. We further summarize the main limitations of current approaches and discuss future research directions.

📄 PDF Abstract BibTeX arXiv:2607.21655

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Survey on Modeling of Human-made Articulated Objects

2024-03-22 · Jiayi Liu, Manolis Savva, Ali Mahdavi-Amiri

3D modeling of articulated objects is a research problem within computer vision, graphics, and robotics. Its objective is to understand the shape and motion of the articulated components, represent the geometry and mobil…

ObjectSurvey

A Survey on Progress in LLM Alignment from the Perspective of Reward Design

2025-05-05 · Miaomiao Ji, Yanqiu Wu, Zhibin Wu, Shoujin Wang 외

The alignment of large language models (LLMs) with human values and intentions represents a core challenge in current AI research, where reward mechanism design has become a critical factor in shaping model behavior. Thi…

TimeRewarder: Learning Dense Reward from Passive Videos via Frame-wise Temporal Distance

2025-09-30 · Yuyang Liu, Chuan Wen, Yihang Hu, Dinesh Jayaraman 외 arxiv

Designing dense rewards is crucial for reinforcement learning (RL), yet in robotics it often demands extensive manual effort and lacks scalability. One promising solution is to view task progress as a dense reward signal…

Reinforcement Learning

Towards a Unified Understanding of Robot Manipulation: A Comprehensive Survey

2025-10-13 · Shuanghao Bai, Wenxuan Song, Jiayi Chen, Yuheng Ji 외 arxiv

Embodied intelligence has witnessed remarkable progress in recent years, driven by advances in computer vision, natural language processing, and the rise of large-scale multimodal models. Among its core challenges, robot…

Robot Manipulation

Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons

2026-03-02 · Anthony Liang, Yigit Korkmaz, Jiahui Zhang, Minyoung Hwang 외 arxiv

General-purpose robot reward models are typically trained to predict absolute task progress from expert demonstrations, providing only local, frame-level supervision. While effective for expert demonstrations, this parad…