paper-with-me

홈 › Papers

Understanding Dynamics of Adam in Zero-Sum Games: An ODE Approach

2026-05-19 · Yi Feng, Weiming Ou, Xiao Wang arxiv

The remarkable success of the Adam in training neural networks has naturally led to the widespread use of its descent-ascent counterpart, Adam-DA, for solving zero-sum games. Despite its popularity in practice, a rigorous theoretical understanding of Adam-DA still lags behind. In this paper, we derive ordinary differential equations (ODEs) that serve as continuous-time limits of the Adam-DA. These ODEs closely approximate the discrete-time dynamics of Adam-DA, providing a tractable analytical framework for understanding its behavior in zero-sum games. Using this ODE approach, we investigate two fundamental aspects of Adam-DA: local convergence and implicit gradient regularization. Our analysis reveals that the roles of the first- and second-order momentum parameters in zero-sum games are exactly the opposite of their well-documented effects in minimization problems. We validate these predictions through GAN experiments across multiple architectures and datasets, demonstrating the practical implications of this reversed momentum effect.

📄 PDF Abstract BibTeX arXiv:2605.19392

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Training GANs with Optimism

2017-10-31 · ICLR 2018 1 · Constantinos Daskalakis, Andrew Ilyas, Vasilis Syrgkanis, Haoyang Zeng

We address the issue of limit cycling behavior in training Generative Adversarial Networks and propose the use of Optimistic Mirror Decent (OMD) for training Wasserstein GANs. Recent theoretical results have shown that o…

Interpreting the Learned Model in MuZero Planning

2024-11-07 · Hung Guei, Yan-Ru Ju, Wei-Yu Chen, Ti-Rong Wu

MuZero has achieved superhuman performance in various games by using a dynamics network to predict environment dynamics for planning, without relying on simulators. However, the latent states learned by the dynamics netw…

Atari GamesBoard Gamesmodel

Recursive Reasoning in Minimax Games: A Level $k$ Gradient Play Method

2022-10-29 · Zichu Liu, Lacra Pavel

Despite the success of generative adversarial networks (GANs) in generating visually appealing images, they are notoriously challenging to train. In order to stabilize the learning dynamics in minimax games, we propose a…

GPUImage GenerationUnconditional Image Generation

Fast Convergence of Optimistic Gradient Ascent in Network Zero-Sum Extensive Form Games

2022-07-18 · Georgios Piliouras, Lillian Ratliff, Ryann Sim, Stratis Skoulakis

The study of learning in games has thus far focused primarily on normal form games. In contrast, our understanding of learning in extensive form games (EFGs) and particularly in EFGs with many agents lags far behind, des…

Form

The Hamiltonian of Poly-matrix Zero-sum Games

2025-05-19 · Toshihiro Ota, Yuma Fujimoto

Understanding a dynamical system fundamentally relies on establishing an appropriate Hamiltonian function and elucidating its symmetries. By formulating agents' strategies and cumulative payoffs as canonically conjugate …