paper-with-me

홈 › Papers

Improved Particle Approximation Error for Mean Field Neural Networks

2024-05-24 · Atsushi Nitanda

Mean-field Langevin dynamics (MFLD) minimizes an entropy-regularized nonlinear convex functional defined over the space of probability distributions. MFLD has gained attention due to its connection with noisy gradient descent for mean-field two-layer neural networks. Unlike standard Langevin dynamics, the nonlinearity of the objective functional induces particle interactions, necessitating multiple particles to approximate the dynamics in a finite-particle setting. Recent works (Chen et al., 2022; Suzuki et al., 2023b) have demonstrated the uniform-in-time propagation of chaos for MFLD, showing that the gap between the particle system and its mean-field limit uniformly shrinks over time as the number of particles increases. In this work, we improve the dependence on logarithmic Sobolev inequality (LSI) constants in their particle approximation errors, which can exponentially deteriorate with the regularization coefficient. Specifically, we establish an LSI-constant-free particle approximation error concerning the objective gap by leveraging the problem structure in risk minimization. As the application, we demonstrate improved convergence of MFLD, sampling guarantee for the mean-field stationary distribution, and uniform-in-time Wasserstein propagation of chaos in terms of particle complexity.

📄 PDF Abstract BibTeX arXiv:2405.15767

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Convergence of mean-field Langevin dynamics: Time and space discretization, stochastic gradient, and variance reduction

2023-06-12 · Taiji Suzuki, Denny Wu, Atsushi Nitanda

The mean-field Langevin dynamics (MFLD) is a nonlinear generalization of the Langevin dynamics that incorporates a distribution-dependent drift, and it naturally arises from the optimization of two-layer neural networks …

Sampling from the Mean-Field Stationary Distribution

2024-02-12 · Yunbum Kook, Matthew S. Zhang, Sinho Chewi, Murat A. Erdogdu 외

We study the complexity of sampling from the stationary distribution of a mean-field SDE, or equivalently, the complexity of minimizing a functional over the space of probability measures which includes an interaction te…

Mean-field Langevin dynamics: Time-space discretization, stochastic gradient, and variance reduction

2023-09-21 · NeurIPS 2023 11

The mean-field Langevin dynamics (MFLD) is a nonlinear generalization of the Langevin dynamics that incorporates a distribution-dependent drift, and it naturally arises from the optimization of two-layer neural networks …

Local exponential stability of mean-field Langevin descent-ascent and associated particle system

2026-02-02 · Geuntaek Seo, Minseop Shin, Pierre Monmarché, Beomjun Choi arxiv

We study the mean-field Langevin descent-ascent (MFL-DA), a coupled optimization dynamics on the space of probability measures for entropically regularized two-player zero-sum games, together with its associated interact…

Convergence of Consensus-Based Particle Methods for Nonconvex Bi-Level Optimization

2026-05-19 · Yutong Chao, Xudong Sun, Konstantin Riedl, Majid Khadiv 외 arxiv

In this paper, we study a consensus-based optimization method for nonconvex bi-level optimization, where the objective is to minimize an upper-level function over the set of global minimizers of a lower-level problem. Th…