paper-with-me

Papers

Learning in games with continuous action sets and unknown payoff functions

2016-08-25 · Panayotis Mertikopoulos, Zhengyuan Zhou

This paper examines the convergence of no-regret learning in games with continuous action sets. For concreteness, we focus on learning via "dual averaging", a widely used class of no-regret learning schemes where players take small steps along their individual payoff gradients and then "mirror" the output back to their action sets. In terms of feedback, we assume that players can only estimate their payoff gradients up to a zero-mean error with bounded variance. To study the convergence of the induced sequence of play, we introduce the notion of variational stability, and we show that stable equilibria are locally attracting with high probability whereas globally stable equilibria are globally attracting with probability 1. We also discuss some applications to mixed-strategy learning in finite games, and we provide explicit estimates of the method's convergence speed.

📄 PDF Abstract BibTeX arXiv:1608.07310

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Undiscounted Bandit Games

2020-08-25

We analyze undiscounted continuous-time games of strategic experimentation with two-armed bandits. The risky arm generates payoffs according to a L\'{e}vy process with an unknown average payoff per unit of time which nat…

State-Constrained Zero-Sum Differential Games with One-Sided Information

2024-03-05 · Mukesh Ghimire, Lei Zhang, Zhe Xu, Yi Ren

We study zero-sum differential games with state constraints and one-sided information, where the informed player (Player 1) has a categorical payoff type unknown to the uninformed player (Player 2). The goal of Player 1 …

A unified stochastic approximation framework for learning in games

2022-06-08 · Panayotis Mertikopoulos, Ya-Ping Hsieh, Volkan Cevher

We develop a flexible stochastic approximation framework for analyzing the long-run behavior of learning in games (both continuous and finite). The proposed analysis template incorporates a wide array of popular learning…

R2-B2: Recursive Reasoning-Based Bayesian Optimization for No-Regret Learning in Games

2020-06-30 · ICML 2020 1 · Zhongxiang Dai, Yizhou Chen, Kian Hsiang Low, Patrick Jaillet 외

This paper presents a recursive reasoning formalism of Bayesian optimization (BO) to model the reasoning process in the interactions between boundedly rational, self-interested agents with unknown, complex, and costly-to…

Bayesian OptimizationMulti-agent Reinforcement Learning

Payoff distribution in robust coalitional games on time-varying networks

2020-10-16

In this paper, we consider a sequence of transferable utility (TU) coalitional games where the coalitional values are unknown but vary within certain bounds. As a solution to the resulting family of games, we formalise t…