paper-with-me

홈 › Papers

Scalable Evaluation of Multi-Agent Reinforcement Learning with Melting Pot

2021-07-14 · Joel Z. Leibo, Edgar Duéñez-Guzmán, Alexander Sasha Vezhnevets, John P. Agapiou, Peter Sunehag, Raphael Koster, Jayd Matyas, Charles Beattie, Igor Mordatch, Thore Graepel

Existing evaluation suites for multi-agent reinforcement learning (MARL) do not assess generalization to novel situations as their primary objective (unlike supervised-learning benchmarks). Our contribution, Melting Pot, is a MARL evaluation suite that fills this gap, and uses reinforcement learning to reduce the human labor required to create novel test scenarios. This works because one agent's behavior constitutes (part of) another agent's environment. To demonstrate scalability, we have created over 80 unique test scenarios covering a broad range of research topics such as social dilemmas, reciprocity, resource sharing, and task partitioning. We apply these test scenarios to standard MARL training algorithms, and demonstrate how Melting Pot reveals weaknesses not apparent from training performance alone.

📄 PDF Abstract BibTeX arXiv:2107.06857

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-agent Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

SocialJax: An Evaluation Suite for Multi-agent Reinforcement Learning in Sequential Social Dilemmas

2025-03-18 · Zihao Guo, Shuqing Shi, Richard Willis, Tristan Tomilin 외

Sequential social dilemmas pose a significant challenge in the field of multi-agent reinforcement learning (MARL), requiring environments that accurately reflect the tension between individual and collective interests. P…

Multi-agent Reinforcement Learningreinforcement-learningReinforcement Learningrllib

Melting Pot 2.0

2022-11-24 · John P. Agapiou, Alexander Sasha Vezhnevets, Edgar A. Duéñez-Guzmán, Jayd Matyas 외

Multi-agent artificial intelligence research promises a path to develop intelligent technologies that are more human-like and more human-compatible than those produced by "solipsistic" approaches, which do not consider i…

Artificial LifeNavigate

Can LLM-Augmented autonomous agents cooperate?, An evaluation of their cooperative capabilities through Melting Pot

2024-03-18 · Manuel Mosquera, Juan Sebastian Pinzon, Manuel Rios, Yesid Fonseca 외

As the field of AI continues to evolve, a significant dimension of this progression is the development of Large Language Models and their potential to enhance multi-agent artificial intelligence systems. This paper explo…

Language ModelingLanguage ModellingLarge Language Model

RPM: Generalizable Behaviors for Multi-Agent Reinforcement Learning

2022-10-18 · Wei Qiu, Xiao Ma, Bo An, Svetlana Obraztsova 외

Despite the recent advancement in multi-agent reinforcement learning (MARL), the MARL agents easily overfit the training environment and perform poorly in the evaluation scenarios where other agents behave differently. O…

Multi-agent Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Beyond the high score: Prosocial ability profiles of multi-agent populations

2025-09-17 · Marko Tesic, Yue Zhao, Joel Z. Leibo, Rakshit S. Trivedi 외 arxiv

The development and evaluation of social capabilities in AI agents require complex environments where competitive and cooperative behaviours naturally emerge. While game-theoretic properties can explain why certain teams…