paper-with-me

Papers

Reactive learning strategies for iterated games

2019-03-11

In an iterated game between two players, there is much interest in characterizing the set of feasible payoffs for both players when one player uses a fixed strategy and the other player is free to switch. Such characterizations have led to extortionists, equalizers, partners, and rivals. Most of those studies use memory-one strategies, which specify the probabilities to take actions depending on the outcome of the previous round. Here, we consider "reactive learning strategies," which gradually modify their propensity to take certain actions based on past actions of the opponent. Every linear reactive learning strategy, $\mathbf{p}^{\ast}$, corresponds to a memory one-strategy, $\mathbf{p}$, and vice versa. We prove that for evaluating the region of feasible payoffs against a memory-one strategy, $\mathcal{C}\left(\mathbf{p}\right)$, we need to check its performance against at most $11$ other strategies. Thus, $\mathcal{C}\left(\mathbf{p}\right)$ is the convex hull in $\mathbb{R}^{2}$ of at most $11$ points. Furthermore, if $\mathbf{p}$ is a memory-one strategy, with feasible payoff region $\mathcal{C}\left(\mathbf{p}\right)$, and $\mathbf{p}^{\ast}$ is the corresponding reactive learning strategy, with feasible payoff region $\mathcal{C}\left(\mathbf{p}^{\ast}\right)$, then $\mathcal{C}\left(\mathbf{p}^{\ast}\right)$ is a subset of $\mathcal{C}\left(\mathbf{p}\right)$. Reactive learning strategies are therefore powerful tools in restricting the outcomes of iterated games.

📄 PDF Abstract BibTeX arXiv:1903.04443

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Autocratic strategies for iterated games with arbitrary action spaces

2016-04-11

The recent discovery of zero-determinant strategies for the iterated Prisoner's Dilemma sparked a surge of interest in the surprising fact that a player can exert unilateral control over iterated interactions. These rema…

Multiagent Reinforcement Learning in Games with an Iterated Dominance Solution

2019-09-25 · Yoram Bachrach, Tor Lattimore, Marta Garnelo, Julien Perolat 외

Multiagent reinforcement learning (MARL) attempts to optimize policies of intelligent agents interacting in the same environment. However, it may fail to converge to a Nash equilibrium in some games. We study independen…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Continuous Adaptation via Meta-Learning in Nonstationary and Competitive Environments

2017-10-10 · ICLR 2018 1 · Maruan Al-Shedivat, Trapit Bansal, Yuri Burda, Ilya Sutskever 외

Ability to continuously learn and adapt from limited experience in nonstationary environments is an important milestone on the path towards general intelligence. In this paper, we cast the problem of continuous adaptatio…

Meta-Learning

Multi-Player Games with LDL Goals over Finite Traces

2020-08-13 · Julian Gutierrez, Giuseppe Perelli, Michael Wooldridge

Linear Dynamic Logic on finite traces LDLf is a powerful logic for reasoning about the behaviour of concurrent and multi-agent systems. In this paper, we investigate techniques for both the characterisation and verificat…

Inferring to C or not to C: Evolutionary games with Bayesian inferential strategies

2023-10-27 · Arunava Patra, Supratim Sengupta, Ayan Paul, Sagar Chakraborty

Strategies for sustaining cooperation and preventing exploitation by selfish agents in repeated games have mostly been restricted to Markovian strategies where the response of an agent depends on the actions in the previ…

Bayesian Inference