paper-with-me

홈 › Papers

Mean-Field Analysis of Two-Layer Neural Networks: Global Optimality with Linear Convergence Rates

2022-05-19 · Jingwei Zhang, Xunpeng Huang

We consider optimizing two-layer neural networks in the mean-field regime where the learning dynamics of network weights can be approximated by the evolution in the space of probability measures over the weight parameters associated with the neurons. The mean-field regime is a theoretically attractive alternative to the NTK (lazy training) regime which is only restricted locally in the so-called neural tangent kernel space around specialized initializations. Several prior works (\cite{chizat2018global, mei2018mean}) establish the asymptotic global optimality of the mean-field regime, but it is still challenging to obtain a quantitative convergence rate due to the complicated unbounded nonlinearity of the training dynamics. This work establishes the first linear convergence result for vanilla two-layer neural networks trained by continuous-time noisy gradient descent in the mean-field regime. Our result relies on a novel time-depdendent estimate of the logarithmic Sobolev constants for a family of measures determined by the evolving distribution of hidden neurons.

📄 PDF Abstract BibTeX arXiv:2205.09860

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

NTK 설명 없음

Similar Papers 제목 키워드 기반

Global optimality of softmax policy gradient with single hidden layer neural networks in the mean-field regime

2020-10-22 · ICLR 2021 1 · Andrea Agazzi, Jianfeng Lu

We study the problem of policy optimization for infinite-horizon discounted Markov Decision Processes with softmax policy and nonlinear function approximation trained with policy gradient algorithms. We concentrate on th…

Global Optimality of Elman-type RNN in the Mean-Field Regime

2023-03-12 · Andrea Agazzi, Jianfeng Lu, Sayan Mukherjee

We analyze Elman-type Recurrent Reural Networks (RNNs) and their training in the mean-field regime. Specifically, we show convergence of gradient descent training dynamics of the RNN to the corresponding mean-field formu…

Vocal Bursts Type Prediction

Global Convergence of Three-layer Neural Networks in the Mean Field Regime

2021-05-11 · ICLR 2021 1 · Huy Tuan Pham, Phan-Minh Nguyen

In the mean field regime, neural networks are appropriately scaled so that as the width tends to infinity, the learning dynamics tends to a nonlinear and nontrivial dynamical limit, known as the mean field limit. This le…

Learning Correlated Equilibria in Mean-Field Games

2022-08-22 · Paul Muller, Romuald Elie, Mark Rowland, Mathieu Lauriere 외

The designs of many large-scale systems today, from traffic routing environments to smart grids, rely on game-theoretic equilibrium concepts. However, as the size of an $N$-player game typically grows exponentially with …

Mean-field Analysis on Two-layer Neural Networks from a Kernel Perspective

2024-03-22 · Shokichi Takakura, Taiji Suzuki

In this paper, we study the feature learning ability of two-layer neural networks in the mean-field regime through the lens of kernel methods. To focus on the dynamics of the kernel induced by the first layer, we utilize…