paper-with-me

홈 › Papers

From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning

2025-01-29 · Junseok Park, Hyeonseo Yang, Min Whoo Lee, Won-Seok Choi, Minsu Lee, Byoung-Tak Zhang

Reinforcement learning (RL) agents often face challenges in balancing exploration and exploitation, particularly in environments where sparse or dense rewards bias learning. Biological systems, such as human toddlers, naturally navigate this balance by transitioning from free exploration with sparse rewards to goal-directed behavior guided by increasingly dense rewards. Inspired by this natural progression, we investigate the Toddler-Inspired Reward Transition in goal-oriented RL tasks. Our study focuses on transitioning from sparse to potential-based dense (S2D) rewards while preserving optimal strategies. Through experiments on dynamic robotic arm manipulation and egocentric 3D navigation tasks, we demonstrate that effective S2D reward transitions significantly enhance learning performance and sample efficiency. Additionally, using a Cross-Density Visualizer, we show that S2D transitions smooth the policy loss landscape, resulting in wider minima that improve generalization in RL models. In addition, we reinterpret Tolman's maze experiments, underscoring the critical role of early free exploratory learning in the context of S2D rewards.

📄 PDF Abstract BibTeX arXiv:2501.17842

Code (0)

등록된 구현이 없습니다.

Tasks

NavigateReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Unveiling the Significance of Toddler-Inspired Reward Transition in Goal-Oriented Reinforcement Learning

2024-03-11 · Junseok Park, Yoonsung Kim, Hee Bin Yoo, Min Whoo Lee 외

Toddlers evolve from free exploration with sparse feedback to exploiting prior experiences for goal-directed learning with denser rewards. Drawing inspiration from this Toddler-Inspired Reward Transition, we set out to e…

Reinforcement Learning (RL)

Toddler-Guidance Learning: Impacts of Critical Period on Multimodal AI Agents

2022-01-12 · Junseok Park, Kwanyoung Park, Hyunseok Oh, Ganghun Lee 외

Critical periods are phases during which a toddler's brain develops in spurts. To promote children's cognitive development, proper guidance is critical in this stage. However, it is not clear whether such a critical peri…

Reinforcement Learning (RL)Transfer Learning

Stage-Transition Dense Reward Modeling for Reinforcement Learning

2026-06-30 · Yang Yang, Bingjie Chen, Zihan Wang, Yizhe Li 외 arxiv

Reinforcement learning for long-horizon robotic manipulation is often limited by sparse and delayed rewards, while manually designing dense shaping signals is costly and brittle to changes in environments and object conf…

Reinforcement Learning

Active Gaze Behavior Boosts Self-Supervised Object Learning

2024-11-04 · Zhengyang Yu, Arthur Aubret, Marcel C. Raabe, Jane Yang 외

Due to significant variations in the projection of the same object from different viewpoints, machine learning algorithms struggle to recognize the same object across various perspectives. In contrast, toddlers quickly l…

ObjectObject RecognitionSelf-Supervised Learning

Learning task-agnostic representation via toddler-inspired learning

2021-01-27 · Kwanyoung Park, Junseok Park, Hyunseok Oh, Byoung-Tak Zhang 외

One of the inherent limitations of current AI systems, stemming from the passive learning mechanisms (e.g., supervised learning), is that they perform well on labeled datasets but cannot deduce knowledge on their own. To…

image-classificationImage ClassificationObject Localization