paper-with-me

Papers

Learning Strategic Value and Cooperation in Multi-Player Stochastic Games through Side Payments

2023-03-09 · Alan Kuhnle, Jeffrey Richley, Darleen Perez-Lavin

For general-sum, n-player, strategic games with transferable utility, the Harsanyi-Shapley value provides a computable method to both 1) quantify the strategic value of a player; and 2) make cooperation rational through side payments. We give a simple formula to compute the HS value in normal-form games. Next, we provide two methods to generalize the HS values to stochastic (or Markov) games, and show that one of them may be computed using generalized Q-learning algorithms. Finally, an empirical validation is performed on stochastic grid-games with three or more players. Source code is provided to compute HS values for both the normal-form and stochastic game setting.

📄 PDF Abstract BibTeX arXiv:2303.05307

Code (0)

등록된 구현이 없습니다.

Tasks

FormQ-Learning

Methods 이 논문이 사용한 방법론

Q-Learning Q-Learning is an off-policy temporal difference control algorithm: $$Q\left(S\_{t}, A\_{t}\right) \leftarrow Q\left(S\_{t}, A\_{t}\right) + \alpha\left[R_{t+1} +…

Similar Papers 제목 키워드 기반

Evolution of Cooperation in Public Goods Games with Stochastic Opting-Out

2017-09-12

This paper investigates the evolution of strategic play where players drawn from a finite well-mixed population are offered the opportunity to play in a public goods game. All players accept the offer. However, due to th…

Multiplayer Bandit Learning, from Competition to Cooperation

2019-08-03 · Simina Brânzei, Yuval Peres

The stochastic multi-armed bandit model captures the tradeoff between exploration and exploitation. We study the effects of competition and cooperation on this tradeoff. Suppose there are $k$ arms and two players, Alice …

Comparing reactive and memory-one strategies of direct reciprocity

2016-05-23

Direct reciprocity is a mechanism for the evolution of cooperation based on repeated interactions. When individuals meet repeatedly, they can use conditional strategies to enforce cooperative outcomes that would not be f…

Noisy information channel mediated prevention of the tragedy of the commons

2024-08-16 · Samrat Sohel Mondal, Sagar Chakraborty

Synergy between evolutionary dynamics of cooperation and fluctuating state of shared resource being consumed by the cooperators is essential for averting the tragedy of the commons. Not only in humans, but also in the co…

Network topology and movement cost, not updating mechanism, determine the evolution of cooperation in mobile structured populations

2023-04-19 · Diogo L. Pires, Igor Erovenko, Mark Broom

Evolutionary models are used to study the self-organisation of collective action, often incorporating population structure due to its ubiquitous presence and long-known impact on emerging phenomena. We investigate the ev…