paper-with-me

홈 › Papers

Multiagent Evaluation under Incomplete Information

2019-09-21 · NeurIPS 2019 12 · Mark Rowland, Shayegan Omidshafiei, Karl Tuyls, Julien Perolat, Michal Valko, Georgios Piliouras, Remi Munos

This paper investigates the evaluation of learned multiagent strategies in the incomplete information setting, which plays a critical role in ranking and training of agents. Traditionally, researchers have relied on Elo ratings for this purpose, with recent works also using methods based on Nash equilibria. Unfortunately, Elo is unable to handle intransitive agent interactions, and other techniques are restricted to zero-sum, two-player settings or are limited by the fact that the Nash equilibrium is intractable to compute. Recently, a ranking method called {\alpha}-Rank, relying on a new graph-based game-theoretic solution concept, was shown to tractably apply to general games. However, evaluations based on Elo or {\alpha}-Rank typically assume noise-free game outcomes, despite the data often being collected from noisy simulations, making this assumption unrealistic in practice. This paper investigates multiagent evaluation in the incomplete information regime, involving general-sum many-player games with noisy outcomes. We derive sample complexity guarantees required to confidently rank agents in this setting. We propose adaptive algorithms for accurate ranking, provide correctness and sample complexity guarantees, then introduce a means of connecting uncertainties in noisy match outcomes to uncertainties in rankings. We evaluate the performance of these approaches in several domains, including Bernoulli games, a soccer meta-game, and Kuhn poker.

📄 PDF Abstract BibTeX arXiv:1909.09849

Code (1)

microsoft/InfoGainalpharank

Similar Papers 제목 키워드 기반

Two-Player Incomplete Games of Resilient Multiagent Systems

2022-12-03 · Yurid Nugraha, Tomohisa Hayakawa, Hideaki Ishii, Ahmet Cetinkaya 외

Evolution of agents' dynamics of multiagent systems under consensus protocol in the face of jamming attacks is discussed, where centralized parties are able to influence the control signals of the agents. In this paper w…

Vocal Bursts Valence Prediction

Comparative Evaluation of Multiagent Learning Algorithms in a Diverse Set of Ad Hoc Team Problems

2019-07-22 · Stefano V. Albrecht, Subramanian Ramamoorthy

This paper is concerned with evaluating different multiagent learning (MAL) algorithms in problems where individual agents may be heterogenous, in the sense of utilizing different learning strategies, without the opportu…

Fairness

JAILJUDGE: A Comprehensive Jailbreak Judge Benchmark with Multi-Agent Enhanced Explanation Evaluation Framework

2024-10-11 · Fan Liu, Yue Feng, Zhao Xu, Lixin Su 외

Despite advancements in enhancing LLM safety against jailbreak attacks, evaluating LLM defenses remains a challenge, with current methods often lacking explainability and generalization to complex scenarios, leading to i…

Networked Multiagent Safe Reinforcement Learning for Low-carbon Demand Management in Distribution Network

2023-11-27 · Jichen Zhang, Linwei Sang, Yinliang Xu, Hongbin Sun

This paper proposes a multiagent based bi-level operation framework for the low-carbon demand management in distribution networks considering the carbon emission allowance on the demand side. In the upper level, the aggr…

ManagementSafe Reinforcement Learning

Estimating counterfactual treatment outcomes over time in complex multiagent scenarios

2022-06-04 · Keisuke Fujii, Koh Takeuchi, Atsushi Kuribayashi, Naoya Takeishi 외

Evaluation of intervention in a multiagent system, e.g., when humans should intervene in autonomous driving systems and when a player should pass to teammates for a good shot, is challenging in various engineering and sc…

Autonomous DrivingcounterfactualPrediction