paper-with-me

Papers

Finite-Time Last-Iterate Convergence for Multi-Agent Learning in Games

2020-02-23 · ICML 2020 1 · Tianyi Lin, Zhengyuan Zhou, Panayotis Mertikopoulos, Michael. I. Jordan

In this paper, we consider multi-agent learning via online gradient descent in a class of games called $\lambda$-cocoercive games, a fairly broad class of games that admits many Nash equilibria and that properly includes unconstrained strongly monotone games. We characterize the finite-time last-iterate convergence rate for joint OGD learning on $\lambda$-cocoercive games; further, building on this result, we develop a fully adaptive OGD learning algorithm that does not require any knowledge of problem parameter (e.g. cocoercive constant $\lambda$) and show, via a novel double-stopping time technique, that this adaptive algorithm achieves same finite-time last-iterate convergence rate as non-adaptive counterpart. Subsequently, we extend OGD learning to the noisy gradient feedback case and establish last-iterate convergence results -- first qualitative almost sure convergence, then quantitative finite-time convergence rates -- all under non-decreasing step-sizes. To our knowledge, we provide the first set of results that fill in several gaps of the existing multi-agent online learning literature, where three aspects -- finite-time convergence rates, non-decreasing step-sizes, and fully adaptive algorithms have been unexplored before.

📄 PDF Abstract BibTeX arXiv:2002.09806

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Last-Iterate Guarantees for Learning in Co-coercive Games

2026-04-21 · Siddharth Chandak, Ramanan Tamizholi, Nicholas Bambos arxiv

We establish finite-time last-iterate guarantees for vanilla stochastic gradient descent in co-coercive games under noisy feedback. This is a broad class of games that is more general than strongly monotone games, allows…

Faster Last-iterate Convergence of Policy Optimization in Zero-Sum Markov Games

2022-10-03 · Shicong Cen, Yuejie Chi, Simon S. Du, Lin Xiao

Multi-Agent Reinforcement Learning (MARL) -- where multiple agents learn to interact in a shared dynamic environment -- permeates across a wide range of critical applications. While there has been substantial progress on…

Multi-agent Reinforcement Learning

Last Iterate Convergence of Incremental Methods and Applications in Continual Learning

2024-03-11 · Xufeng Cai, Jelena Diakonikolas

Incremental gradient and incremental proximal methods are a fundamental class of optimization algorithms used for solving finite sum problems, broadly studied in the literature. Yet, without strong convexity, their conve…

Continual Learning

Augmented Lagrangian Method for Last-Iterate Convergence for Constrained MDPs

2026-05-12 · Michael Lu, Max Qiushi Lin, Mo Chen, Sharan Vaswani arxiv

We study policy optimization for infinite-horizon, discounted constrained Markov decision processes (CMDPs). While existing theoretical guarantees typically hold for the mixture policy, deploying such a policy is computa…

Continuous Control

On Separation Between Best-Iterate, Random-Iterate, and Last-Iterate Convergence of Learning in Games

2025-03-04 · Yang Cai, Gabriele Farina, Julien Grand-Clément, Christian Kroer 외

Non-ergodic convergence of learning dynamics in games is widely studied recently because of its importance in both theory and practice. Recent work (Cai et al., 2024) showed that a broad class of learning dynamics, inclu…