paper-with-me

홈 › Papers

Mean-field theory of two-layers neural networks: dimension-free bounds and kernel limit

2019-02-16 · Song Mei, Theodor Misiakiewicz, Andrea Montanari

We consider learning two layer neural networks using stochastic gradient descent. The mean-field description of this learning dynamics approximates the evolution of the network weights by an evolution in the space of probability distributions in $R^D$ (where $D$ is the number of parameters associated to each neuron). This evolution can be defined through a partial differential equation or, equivalently, as the gradient flow in the Wasserstein space of probability distributions. Earlier work shows that (under some regularity assumptions), the mean field description is accurate as soon as the number of hidden units is much larger than the dimension $D$. In this paper we establish stronger and more general approximation guarantees. First of all, we show that the number of hidden units only needs to be larger than a quantity dependent on the regularity properties of the data, and independent of the dimensions. Next, we generalize this analysis to the case of unbounded activation functions, which was not covered by earlier bounds. We extend our results to noisy stochastic gradient descent. Finally, we show that kernel ridge regression can be recovered as a special limit of the mean field analysis.

📄 PDF Abstract BibTeX arXiv:1902.06015

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Deep Learning for Mean Field Games and Mean Field Control with Applications to Finance

2021-07-09 · René Carmona, Mathieu Laurière

Financial markets and more generally macro-economic models involve a large number of individuals interacting through variables such as prices resulting from the aggregate behavior of all the agents. Mean field games have…

Federated Learning as a Mean-Field Game

2021-07-08 · Arash Mehrjou

We establish a connection between federated learning, a concept from machine learning, and mean-field games, a concept from game theory and control theory. In this analogy, the local federated learners are considered as …

Federated LearningPrivacy Preserving

Infinite Limits of Multi-head Transformer Dynamics

2024-05-24 · Blake Bordelon, Hamza Tahir Chaudhry, Cengiz Pehlevan

In this work, we analyze various scaling limits of the training dynamics of transformer models in the feature learning regime. We identify the set of parameterizations that admit well-defined infinite width and depth lim…

Game-theoretical control with continuous action sets

2014-12-01 · Steven Perkins, Panayotis Mertikopoulos, David S. Leslie

Motivated by the recent applications of game-theoretical learning techniques to the design of distributed control systems, we study a class of control problems that can be formulated as potential games with continuous ac…

Reinforcement Learning

Polynomial-time Sparse Measure Recovery: From Mean Field Theory to Algorithm Design

2022-04-16 · Hadi Daneshmand, Francis Bach

Mean field theory has provided theoretical insights into various algorithms by letting the problem size tend to infinity. We argue that the applications of mean-field theory go beyond theoretical insights as it can inspi…

Super-ResolutionTensor Decomposition