paper-with-me

홈 › Papers

Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals

2025-07-02 · Yannick Molinghen, Tom Lenaerts arxiv

This work re-examines the commonly held assumption that the frequency of rewards is a reliable measure of task difficulty in reinforcement learning. We identify and formalize a structural challenge that undermines the effectiveness of current policy learning methods: when essential subgoals do not directly yield rewards. We characterize such settings as exhibiting zero-incentive dynamics, where transitions critical to success remain unrewarded. We show that state-of-the-art deep subgoal-based algorithms fail to leverage these dynamics and that learning performance is highly sensitive to the temporal proximity between subgoal completion and eventual reward. These findings reveal a fundamental limitation in current approaches and point to the need for mechanisms that can infer latent task structure without relying on immediate incentives.

📄 PDF Abstract BibTeX arXiv:2507.01470

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Cost optimisation of individual-based institutional reward incentives for promoting cooperation in finite populations

2024-02-12 · M. H. Duong, C. M. Durbac, T. A. Han

In this paper, we study the problem of cost optimisation of individual-based institutional incentives (reward, punishment, and hybrid) for guaranteeing a certain minimal level of cooperative behaviour in a well-mixed, fi…

Laser Learning Environment: A new environment for coordination-critical multi-agent tasks

2024-04-04 · Yannick Molinghen, Raphaël Avalos, Mark Van Achter, Ann Nowé 외

We introduce the Laser Learning Environment (LLE), a collaborative multi-agent reinforcement learning environment in which coordination is central. In LLE, agents depend on each other to make progress (interdependence), …

Multi-agent Reinforcement LearningQ-Learning

Social welfare optimisation under institutional reward and punishment

2026-05-29 · Van An Nguyen, Vuong Khang Huynh, Huu Loi Bui, Hai Anh Ha 외 arxiv

Institutional incentives are widely used to promote cooperation among autonomous, self-regarding agents, from human societies to multi-agent and AI systems. Existing work typically treats incentive design as a bi-objecti…

Offsetting Unequal Competition through RL-assisted Incentive Schemes

2022-01-05 · Paramita Koley, Aurghya Maiti, Sourangshu Bhattacharya, Niloy Ganguly

This paper investigates the dynamics of competition among organizations with unequal expertise. Multi-agent reinforcement learning has been used to simulate and understand the impact of various incentive schemes designed…

Multi-agent Reinforcement Learningreinforcement-learningReinforcement Learning (RL)

NOVER: Incentive Training for Language Models via Verifier-Free Reinforcement Learning

2025-05-21 · Wei Liu, Siya Qi, Xinyu Wang, Chen Qian 외

Recent advances such as DeepSeek R1-Zero highlight the effectiveness of incentive training, a reinforcement learning paradigm that computes rewards solely based on the final answer part of a language model's output, ther…

General Reinforcement LearningLogical ReasoningOpen-Domain Question Answeringreinforcement-learning+1