Some Considerations on Learning to Explore via Meta-Reinforcement Learning
We consider the problem of exploration in meta reinforcement learning. Two new meta reinforcement learning algorithms are suggested: E-MAML and E-$\text{RL}^2$. Results are presented on a novel environment we call `Krazy World' and a set of maze environments. We show E-MAML and E-$\text{RL}^2$ deliver better performance on tasks where exploration is important.
Code (7)
Tasks
Meta Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Meta-learning Extractors for Music Source Separation
We propose a hierarchical meta-learning-inspired model for music source separation (Meta-TasNet) in which a generator model is used to predict the weights of individual extractor models. This enables efficient parameter-…
Meta-LearningMusic Source SeparationCan Meta-Interpretive Learning outperform Deep Reinforcement Learning of Evaluable Game strategies?
World-class human players have been outperformed in a number of complex two person games (Go, Chess, Checkers) by Deep Reinforcement Learning systems. However, owing to tractability considerations minimax regret of a lea…
Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1Graph Reinforcement Learning for Operator Selection in the ALNS Metaheuristic
ALNS is a popular metaheuristic with renowned efficiency in solving combinatorial optimisation problems. However, despite 16 years of intensive research into ALNS, whether the embedded adaptive layer can efficiently sele…
Deep Reinforcement LearningOpen-Ended Question Answeringreinforcement-learningReinforcement Learning+1Generative methods for sampling transition paths in molecular dynamics
Molecular systems often remain trapped for long times around some local minimum of the potential energy function, before switching to another one -- a behavior known as metastability. Simulating transition paths linking …
reinforcement-learningReinforcement Learning (RL)Composing Meta-Policies for Autonomous Driving Using Hierarchical Deep Reinforcement Learning
Rather than learning new control policies for each new task, it is possible, when tasks share some structure, to compose a "meta-policy" from previously learned policies. This paper reports results from experiments using…
Autonomous DrivingDeep Reinforcement Learningreinforcement-learningReinforcement Learning+1