paper-with-me

홈 › Papers

Retrospective on the 2021 BASALT Competition on Learning from Human Feedback

2022-04-14 · Rohin Shah, Steven H. Wang, Cody Wild, Stephanie Milani, Anssi Kanervisto, Vinicius G. Goecks, Nicholas Waytowich, David Watkins-Valls, Bharat Prakash, Edmund Mills, Divyansh Garg, Alexander Fries, Alexandra Souly, Chan Jun Shern, Daniel del Castillo, Tom Lieberum

We held the first-ever MineRL Benchmark for Agents that Solve Almost-Lifelike Tasks (MineRL BASALT) Competition at the Thirty-fifth Conference on Neural Information Processing Systems (NeurIPS 2021). The goal of the competition was to promote research towards agents that use learning from human feedback (LfHF) techniques to solve open-world tasks. Rather than mandating the use of LfHF techniques, we described four tasks in natural language to be accomplished in the video game Minecraft, and allowed participants to use any approach they wanted to build agents that could accomplish the tasks. Teams developed a diverse range of LfHF algorithms across a variety of possible human feedback types. The three winning teams implemented significantly different approaches while achieving similar performance. Interestingly, their approaches performed well on different tasks, validating our choice of tasks to include in the competition. While the outcomes validated the design of our competition, we did not get as many participants and submissions as our sister competition, MineRL Diamond. We speculate about the causes of this problem and suggest improvements for future iterations of the competition.

📄 PDF Abstract BibTeX arXiv:2204.07123

Code (0)

등록된 구현이 없습니다.

Tasks

Minecraft

Similar Papers 제목 키워드 기반

Towards Solving Fuzzy Tasks with Human Feedback: A Retrospective of the MineRL BASALT 2022 Competition

2023-03-23 · Stephanie Milani, Anssi Kanervisto, Karolis Ramanauskas, Sander Schulhoff 외

To facilitate research in the direction of fine-tuning foundation models from human feedback, we held the MineRL BASALT Competition on Fine-Tuning from Human Feedback at NeurIPS 2022. The BASALT challenge asks teams to c…

Minecraft

BEDD: The MineRL BASALT Evaluation and Demonstrations Dataset for Training and Benchmarking Agents that Solve Fuzzy Tasks

2023-12-05 · NeurIPS 2023 11 · Stephanie Milani, Anssi Kanervisto, Karolis Ramanauskas, Sander Schulhoff 외

The MineRL BASALT competition has served to catalyze advances in learning from human feedback through four hard-to-specify tasks in Minecraft, such as create and photograph a waterfall. Given the completion of two years …

BenchmarkingMinecraft

Combining Learning from Human Feedback and Knowledge Engineering to Solve Hierarchical Tasks in Minecraft

2021-12-07 · Vinicius G. Goecks, Nicholas Waytowich, David Watkins-Valls, Bharat Prakash

Real-world tasks of interest are generally poorly defined by human-readable descriptions and have no pre-defined reward signals unless it is defined by a human designer. Conversely, data-driven algorithms are often desig…

Imitation LearningMinecraft

The MineRL BASALT Competition on Learning from Human Feedback

2021-07-05 · Rohin Shah, Cody Wild, Steven H. Wang, Neel Alex 외

The last decade has seen a significant increase of interest in deep learning research, with many public successes that have demonstrated its potential. As such, these systems are now being incorporated into commercial pr…

Imitation LearningMinecraft

DIP-RL: Demonstration-Inferred Preference Learning in Minecraft

2023-07-22 · Ellen Novoseller, Vinicius G. Goecks, David Watkins, Josh Miller 외

In machine learning for sequential decision-making, an algorithmic agent learns to interact with an environment while receiving feedback in the form of a reward signal. However, in many unstructured real-world settings, …

Decision MakingMinecraftreinforcement-learningReinforcement Learning+2