paper-with-me

Papers

A Fisher-Rao gradient flow for entropic mean-field min-max games

2024-05-24 · Razvan-Andrei Lascu, Mateusz B. Majka, Łukasz Szpruch

Gradient flows play a substantial role in addressing many machine learning problems. We examine the convergence in continuous-time of a \textit{Fisher-Rao} (Mean-Field Birth-Death) gradient flow in the context of solving convex-concave min-max games with entropy regularization. We propose appropriate Lyapunov functions to demonstrate convergence with explicit rates to the unique mixed Nash equilibrium.

📄 PDF Abstract BibTeX arXiv:2405.15834

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

On propagation of chaos for the Fisher-Rao gradient flow in entropic mean-field optimization

2026-02-16 · Petra Lazić, Linshan Liu, Mateusz B. Majka arxiv

We consider a class of optimization problems on the space of probability measures motivated by the mean-field approach to studying neural networks. Such problems can be solved by constructing continuous-time gradient flo…

Fisher-Rao Gradient Flows of Linear Programs and State-Action Natural Policy Gradients

2024-03-28 · Johannes Müller, Semih Çaycı, Guido Montúfar

Kakade's natural policy gradient method has been studied extensively in recent years, showing linear convergence with and without regularization. We study another natural gradient method based on the Fisher information m…

Sampling in Unit Time with Kernel Fisher-Rao Flow

2024-01-08 · Aimee Maurais, Youssef Marzouk

We introduce a new mean-field ODE and corresponding interacting particle systems (IPS) for sampling from an unnormalized target density. The IPS are gradient-free, available in closed form, and only require the ability t…

Convergence of Policy Gradient for Entropy Regularized MDPs with Neural Network Approximation in the Mean-Field Regime

2022-01-18 · Bekzhan Kerimkulov, James-Michael Leahy, David Šiška, Lukasz Szpruch

We study the global convergence of policy gradient for infinite-horizon, continuous state and action space, and entropy-regularized Markov decision processes (MDPs). We consider a softmax policy with (one-hidden layer) n…

Mean Field Optimization Problem Regularized by Fisher Information

2023-02-12 · Julien Claisse, Giovanni Conforti, Zhenjie Ren, SongBo Wang

Recently there is a rising interest in the research of mean field optimization, in particular because of its role in analyzing the training of neural networks. In this paper by adding the Fisher Information as the regula…