paper-with-me

홈 › Papers

VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning

2019-10-18 · ICLR 2020 1 · Luisa Zintgraf, Kyriacos Shiarlis, Maximilian Igl, Sebastian Schulze, Yarin Gal, Katja Hofmann, Shimon Whiteson

Trading off exploration and exploitation in an unknown environment is key to maximising expected return during learning. A Bayes-optimal policy, which does so optimally, conditions its actions not only on the environment state but on the agent's uncertainty about the environment. Computing a Bayes-optimal policy is however intractable for all but the smallest tasks. In this paper, we introduce variational Bayes-Adaptive Deep RL (variBAD), a way to meta-learn to perform approximate inference in an unknown environment, and incorporate task uncertainty directly during action selection. In a grid-world domain, we illustrate how variBAD performs structured online exploration as a function of task uncertainty. We further evaluate variBAD on MuJoCo domains widely used in meta-RL and show that it achieves higher online return than existing methods.

📄 PDF Abstract BibTeX arXiv:1910.08348

Code (3)

lmzintgraf/varibad 공식 구현 pytorch
ido90/robustmetarl
twni2016/pomdp-baselines pytorch

Tasks

Meta-LearningMuJoCo

Similar Papers 제목 키워드 기반

Offline Meta Learning of Exploration

2020-08-06 · NeurIPS 2021 12 · Ron Dorfman, Idan Shenfeld, Aviv Tamar

Consider the following instance of the Offline Meta Reinforcement Learning (OMRL) problem: given the complete training logs of $N$ conventional RL agents, trained on $N$ different tasks, design a meta-agent that can quic…

Meta-LearningMeta Reinforcement Learning

Offline Meta Reinforcement Learning -- Identifiability Challenges and Effective Data Collection Strategies

2021-05-21 · NeurIPS 2021 12 · Ron Dorfman, Idan Shenfeld, Aviv Tamar

Consider the following instance of the Offline Meta Reinforcement Learning (OMRL) problem: given the complete training logs of $N$ conventional RL agents, trained on $N$ different tasks, design a meta-agent that can quic…

Meta Reinforcement Learningreinforcement-learningReinforcement Learning (RL)

Meta Reinforcement Learning with Finite Training Tasks -- a Density Estimation Approach

2022-06-21 · Zohar Rimon, Aviv Tamar, Gilad Adler

In meta reinforcement learning (meta RL), an agent learns from a set of training tasks how to quickly solve a new task, drawn from the same task distribution. The optimal meta RL policy, a.k.a. the Bayes-optimal behavior…

Density EstimationDimensionality ReductionMeta Reinforcement Learningreinforcement-learning+1

Bayesian Matrix Completion via Adaptive Relaxed Spectral Regularization

2015-12-03 · Yang Song, Jun Zhu

Bayesian matrix completion has been studied based on a low-rank matrix factorization formulation with promising results. However, little work has been done on Bayesian matrix completion based on the more direct spectral …

Bayesian InferenceCollaborative FilteringMatrix Completion

Amortized Bayesian Meta-Learning

2019-05-01 · ICLR 2019 5 · Sachin Ravi, Alex Beatson

Meta-learning, or learning-to-learn, has proven to be a successful strategy in attacking problems in supervised learning and reinforcement learning that involve small amounts of data. State-of-the-art solutions involve l…

Few-Shot Image ClassificationFew-Shot LearningMeta-LearningReinforcement Learning+1