paper-with-me

홈 › Papers

Bottom-Up Meta-Policy Search

2019-10-22 · Luckeciano C. Melo, Marcos R. O. A. Maximo, Adilson Marques da Cunha

Despite of the recent progress in agents that learn through interaction, there are several challenges in terms of sample efficiency and generalization across unseen behaviors during training. To mitigate these problems, we propose and apply a first-order Meta-Learning algorithm called Bottom-Up Meta-Policy Search (BUMPS), which works with two-phase optimization procedure: firstly, in a meta-training phase, it distills few expert policies to create a meta-policy capable of generalizing knowledge to unseen tasks during training; secondly, it applies a fast adaptation strategy named Policy Filtering, which evaluates few policies sampled from the meta-policy distribution and selects which best solves the task. We conducted all experiments in the RoboCup 3D Soccer Simulation domain, in the context of kick motion learning. We show that, given our experimental setup, BUMPS works in scenarios where simple multi-task Reinforcement Learning does not. Finally, we performed experiments in a way to evaluate each component of the algorithm.

📄 PDF Abstract BibTeX arXiv:1910.10232

Code (1)

luckeciano/bumps 공식 구현 tf

Tasks

Meta-LearningReinforcement Learning

Similar Papers 제목 키워드 기반

Unifying Top-down and Bottom-up for Recurrent Visual Attention

2021-09-29 · Gang Chen

The idea of using the recurrent neural network for visual attention has gained popularity in computer vision community. Although the recurrent visual attention model (RAM) leverages the glimpses with more large patch siz…

Q-Learning

CrossBeam: Learning to Search in Bottom-Up Program Synthesis

2022-03-20 · ICLR 2022 4 · Kensen Shi, Hanjun Dai, Kevin Ellis, Charles Sutton

Many approaches to program synthesis perform a search within an enormous space of programs to find one that satisfies a given specification. Prior works have used neural models to guide combinatorial search algorithms, b…

Program SynthesisStructured Prediction

Where to Look: A Unified Attention Model for Visual Recognition with Reinforcement Learning

2021-11-13 · Gang Chen

The idea of using the recurrent neural network for visual attention has gained popularity in computer vision community. Although the recurrent attention model (RAM) leverages the glimpses with more large patch size to in…

Q-LearningReinforcement Learning (RL)

Contextualizing Artificially Intelligent Morality: A Meta-Ethnography of Top-Down, Bottom-Up, and Hybrid Models for Theoretical and Applied Ethics in Artificial Intelligence

2022-04-15 · Jennafer S. Roberts, Laura N. Montoya

In this meta-ethnography, we explore three different angles of ethical artificial intelligence (AI) design implementation including the philosophical ethical viewpoint, the technical perspective, and framing through a po…

Ethics

Annotating a Russian corpus of conceptual metaphor: a bottom-up approach

2013-06-01 · WS 2013 6 · Yulia Badryzlova, Natalia Shekhtman, Yekaterina Isaeva, Ruslan Kerimov