paper-with-me

Papers

Retrospective Analysis of the 2019 MineRL Competition on Sample Efficient Reinforcement Learning

2020-03-10 · Stephanie Milani, Nicholay Topin, Brandon Houghton, William H. Guss, Sharada P. Mohanty, Keisuke Nakata, Oriol Vinyals, Noboru Sean Kuno

To facilitate research in the direction of sample efficient reinforcement learning, we held the MineRL Competition on Sample Efficient Reinforcement Learning Using Human Priors at the Thirty-third Conference on Neural Information Processing Systems (NeurIPS 2019). The primary goal of this competition was to promote the development of algorithms that use human demonstrations alongside reinforcement learning to reduce the number of samples needed to solve complex, hierarchical, and sparse environments. We describe the competition, outlining the primary challenge, the competition design, and the resources that we provided to the participants. We provide an overview of the top solutions, each of which use deep reinforcement learning and/or imitation learning. We also discuss the impact of our organizational decisions on the competition and future directions for improvement.

📄 PDF Abstract BibTeX arXiv:2003.05012

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement LearningImitation Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Towards Solving Fuzzy Tasks with Human Feedback: A Retrospective of the MineRL BASALT 2022 Competition

2023-03-23 · Stephanie Milani, Anssi Kanervisto, Karolis Ramanauskas, Sander Schulhoff 외

To facilitate research in the direction of fine-tuning foundation models from human feedback, we held the MineRL BASALT Competition on Fine-Tuning from Human Feedback at NeurIPS 2022. The BASALT challenge asks teams to c…

Minecraft

SEIHAI: A Sample-efficient Hierarchical AI for the MineRL Competition

2021-11-17 · Hangyu Mao, Chao Wang, Xiaotian Hao, Yihuan Mao 외

The MineRL competition is designed for the development of reinforcement learning and imitation learning algorithms that can efficiently leverage human demonstrations to drastically reduce the number of environment intera…

Imitation Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Retrospective on the 2021 BASALT Competition on Learning from Human Feedback

2022-04-14 · Rohin Shah, Steven H. Wang, Cody Wild, Stephanie Milani 외

We held the first-ever MineRL Benchmark for Agents that Solve Almost-Lifelike Tasks (MineRL BASALT) Competition at the Thirty-fifth Conference on Neural Information Processing Systems (NeurIPS 2021). The goal of the comp…

Minecraft

The MineRL 2020 Competition on Sample Efficient Reinforcement Learning using Human Priors

2021-01-26 · William H. Guss, Mario Ynocente Castro, Sam Devlin, Brandon Houghton 외

Although deep reinforcement learning has led to breakthroughs in many difficult domains, these successes have required an ever-increasing number of samples, affording only a shrinking segment of the AI community access t…

Decision MakingDeep Reinforcement LearningEfficient ExplorationMinecraft+3

The MineRL 2019 Competition on Sample Efficient Reinforcement Learning using Human Priors

2019-04-22 · William H. Guss, Cayden Codel, Katja Hofmann, Brandon Houghton 외

Though deep reinforcement learning has led to breakthroughs in many difficult domains, these successes have required an ever-increasing number of samples. As state-of-the-art reinforcement learning (RL) systems require a…

Decision MakingDeep Reinforcement LearningEfficient ExplorationMinecraft+4