paper-with-me

Papers

Countering Feedback Delays in Multi-Agent Learning

2017-12-01 · NeurIPS 2017 12 · Zhengyuan Zhou, Panayotis Mertikopoulos, Nicholas Bambos, Peter W. Glynn, Claire Tomlin

We consider a model of game-theoretic learning based on online mirror descent (OMD) with asynchronous and delayed feedback information. Instead of focusing on specific games, we consider a broad class of continuous games defined by the general equilibrium stability notion, which we call λ-variational stability. Our first contribution is that, in this class of games, the actual sequence of play induced by OMD-based learning converges to Nash equilibria provided that the feedback delays faced by the players are synchronous and bounded. Subsequently, to tackle fully decentralized, asynchronous environments with (possibly) unbounded delays between actions and feedback, we propose a variant of OMD which we call delayed mirror descent (DMD), and which relies on the repeated leveraging of past information. With this modification, the algorithm converges to Nash equilibria with no feedback synchronicity assumptions and even when the delays grow superlinearly relative to the horizon of play.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Multi-Agent Online Optimization with Delays: Asynchronicity, Adaptivity, and Optimism

2020-12-21 · Yu-Guan Hsieh, Franck Iutzeler, Jérôme Malick, Panayotis Mertikopoulos

In this paper, we provide a general framework for studying multi-agent online learning problems in the presence of delays and asynchronicities. Specifically, we propose and analyze a class of adaptive dual averaging sche…

Revisiting Multi-Agent Asynchronous Online Optimization with Delays: the Strongly Convex Case

2025-03-13 · Lingchan Bao, Tong Wei, Yuanyu Wan

We revisit multi-agent asynchronous online optimization with delays, where only one of the agents becomes active for making the decision at each round, and the corresponding feedback is received by all the agents after u…

A Reduction-based Framework for Sequential Decision Making with Delayed Feedback

2023-02-03 · NeurIPS 2023 11 · Yunchang Yang, Han Zhong, Tianhao Wu, Bin Liu 외

We study stochastic delayed feedback in general multi-agent sequential decision making, which includes bandits, single-agent Markov decision processes (MDPs), and Markov games (MGs). We propose a novel reduction-based fr…

Decision MakingSequential Decision Making

Decentralized Online Convex Optimization with Unknown Feedback Delays

2026-01-12 · Hao Qiu, Mengxiao Zhang, Juliette Achddou arxiv

Decentralized online convex optimization (D-OCO), where multiple agents within a network collaboratively learn optimal decisions in real-time, arises naturally in applications such as federated learning, sensor networks,…

Federated Learning

Asynchronous Gradient Play in Zero-Sum Multi-agent Games

2022-11-16 · Ruicheng Ao, Shicong Cen, Yuejie Chi

Finding equilibria via gradient play in competitive multi-agent games has been attracting a growing amount of attention in recent years, with emphasis on designing efficient strategies where the agents operate in a decen…