paper-with-me

홈 › Papers

Hindsight Task Relabelling: Experience Replay for Sparse Reward Meta-RL

2021-12-02 · NeurIPS 2021 12 · Charles Packer, Pieter Abbeel, Joseph E. Gonzalez

Meta-reinforcement learning (meta-RL) has proven to be a successful framework for leveraging experience from prior tasks to rapidly learn new related tasks, however, current meta-RL approaches struggle to learn in sparse reward environments. Although existing meta-RL algorithms can learn strategies for adapting to new sparse reward tasks, the actual adaptation strategies are learned using hand-shaped reward functions, or require simple environments where random exploration is sufficient to encounter sparse reward. In this paper, we present a formulation of hindsight relabeling for meta-RL, which relabels experience during meta-training to enable learning to learn entirely using sparse reward. We demonstrate the effectiveness of our approach on a suite of challenging sparse reward goal-reaching environments that previously required dense reward during meta-training to solve. Our approach solves these environments using the true sparse reward function, with performance comparable to training with a proxy dense reward function.

📄 PDF Abstract BibTeX arXiv:2112.00901

Code (0)

등록된 구현이 없습니다.

Tasks

Meta Reinforcement Learning

Similar Papers 제목 키워드 기반

Diversity-based Trajectory and Goal Selection with Hindsight Experience Replay

2021-08-17 · Tianhong Dai, Hengyan Liu, Kai Arulkumaran, Guangyu Ren 외

Hindsight experience replay (HER) is a goal relabelling technique typically used with off-policy deep reinforcement learning algorithms to solve goal-oriented tasks; it is well suited to robotic manipulation tasks that d…

Deep Reinforcement LearningDiversityPoint Processes

Hindsight Curriculum Generation Based Multi-Goal Experience Replay

2021-01-01 · Xiaoyun Feng

In multi-goal tasks with sparse rewards, it is challenging to learn from tons of experiences with zero rewards. Hindsight experience replay (HER), which replays past experiences with additional heuristic goals, has shown…

Reinforcement Learning (RL)

SPLID: Self-Imitation Policy Learning through Iterative Distillation

2021-09-29 · Zhihan Liu, Hao Sun, Bolei Zhou

Goal-Conditioned continuous control tasks remain challenging due to the sparse reward signals. To address this issue, many relabelling methods like Hindsight Experience Replay have been developed and bring significant im…

continuous-controlContinuous Control

Improvements on Hindsight Learning

2018-09-16 · Ameet Deshpande, Srikanth Sarma, Ashutosh Jha, Balaraman Ravindran

Sparse reward problems are one of the biggest challenges in Reinforcement Learning. Goal-directed tasks are one such sparse reward problems where a reward signal is received only when the goal is reached. One promising w…

Policy Gradient Methodsreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Imaginary Hindsight Experience Replay: Curious Model-based Learning for Sparse Reward Tasks

2021-10-05 · Robert McCarthy, Qiang Wang, Stephen J. Redmond

Model-based reinforcement learning is a promising learning strategy for practical robotic applications due to its improved data-efficiency versus model-free counterparts. However, current state-of-the-art model-based met…

FetchPush-v1Model-based Reinforcement LearningOpenAI Gym