paper-with-me

홈 › Papers

Model Ensemble-Based Intrinsic Reward for Sparse Reward Reinforcement Learning

2019-09-25 · Giseung Park, Whiyoung Jung, Sungho Choi, Youngchul Sung

In this paper, a new intrinsic reward generation method for sparse-reward reinforcement learning is proposed based on an ensemble of dynamics models. In the proposed method, the mixture of multiple dynamics models is used to approximate the true unknown transition probability, and the intrinsic reward is designed as the minimum of the surprise seen from each dynamics model to the mixture of the dynamics models. In order to show the effectiveness of the proposed intrinsic reward generation method, a working algorithm is constructed by combining the proposed intrinsic reward generation method with the proximal policy optimization (PPO) algorithm. Numerical results show that for representative locomotion tasks, the proposed model-ensemble-based intrinsic reward generation method outperforms the previous methods based on a single dynamics model.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Curiosity-driven Exploration in Sparse-reward Multi-agent Reinforcement Learning

2023-02-21 · Jiong Li, Pratik Gajane

Sparsity of rewards while applying a deep reinforcement learning method negatively affects its sample-efficiency. A viable solution to deal with the sparsity of rewards is to learn via intrinsic motivation which advocate…

Deep Reinforcement LearningMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning+1

Successor-Predecessor Intrinsic Exploration

2023-05-24 · NeurIPS 2023 11 · Changmin Yu, Neil Burgess, Maneesh Sahani, Samuel J. Gershman

Exploration is essential in reinforcement learning, particularly in environments where external rewards are sparse. Here we focus on exploration with intrinsic rewards, where the agent transiently augments the external r…

Atari GamesDeep Reinforcement LearningEfficient Explorationreinforcement-learning+1

Adaptive Multi-model Fusion Learning for Sparse-Reward Reinforcement Learning

2021-01-01 · Giseung Park, Whiyoung Jung, Sungho Choi, Youngchul Sung

In this paper, we consider intrinsic reward generation for sparse-reward reinforcement learning based on model prediction errors. In typical model-prediction-error-based intrinsic reward generation, an agent has a learni…

Predictionreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Adaptive Correlation-Weighted Intrinsic Rewards for Reinforcement Learning

2026-02-27 · Viet Bac Nguyen, Phuong Thai Nguyen arxiv

We propose ACWI (Adaptive Correlation Weighted Intrinsic), an adaptive intrinsic reward scaling framework designed to dynamically balance intrinsic and extrinsic rewards for improved exploration in sparse reward reinforc…

Computational EfficiencyReinforcement Learning

Implicit Generative Modeling for Efficient Exploration

2019-11-19 · ICML 2020 1 · Neale Ratzlaff, Qinxun Bai, Li Fuxin, Wei Xu

Efficient exploration remains a challenging problem in reinforcement learning, especially for those tasks where rewards from environments are sparse. A commonly used approach for exploring such environments is to introdu…

Efficient ExplorationFuture predictionReinforcement Learning