paper-with-me

홈 › Papers

A Note on the Global Convergence of Multilayer Neural Networks in the Mean Field Regime

2020-06-16 · Huy Tuan Pham, Phan-Minh Nguyen

In a recent work, we introduced a rigorous framework to describe the mean field limit of the gradient-based learning dynamics of multilayer neural networks, based on the idea of a neuronal embedding. There we also proved a global convergence guarantee for three-layer (as well as two-layer) networks using this framework. In this companion note, we point out that the insights in our previous work can be readily extended to prove a global convergence guarantee for multilayer networks of any depths. Unlike our previous three-layer global convergence guarantee that assumes i.i.d. initializations, our present result applies to a type of correlated initialization. This initialization allows to, at any finite training time, propagate a certain universal approximation property through the depth of the neural network. To achieve this effect, we introduce a bidirectional diversity condition.

📄 PDF Abstract BibTeX arXiv:2006.09355

Code (0)

등록된 구현이 없습니다.

Tasks

Diversity

Similar Papers 제목 키워드 기반

A Rigorous Framework for the Mean Field Limit of Multilayer Neural Networks

2020-01-30 · Phan-Minh Nguyen, Huy Tuan Pham

We develop a mathematically rigorous framework for multilayer neural networks in the mean field regime. As the network's widths increase, the network's learning trajectory is shown to be well captured by a meaningful and…

Global Convergence of Three-layer Neural Networks in the Mean Field Regime

2021-05-11 · ICLR 2021 1 · Huy Tuan Pham, Phan-Minh Nguyen

In the mean field regime, neural networks are appropriately scaled so that as the width tends to infinity, the learning dynamics tends to a nonlinear and nontrivial dynamical limit, known as the mean field limit. This le…

A note on convergence of Wasserstein policy optimization

2026-05-21 · David Šiška, Yufei Zhang arxiv

Wasserstein Policy Optimization (WPO) is a recently proposed reinforcement learning algorithm that leverages Wasserstein gradient flows to optimize stochastic policies in continuous action spaces. Despite its empirical s…

Reinforcement Learning

A Mean-field Analysis of Deep ResNet and Beyond: Towards Provable Optimization Via Overparameterization From Depth

2020-03-11 · Yiping Lu, Chao Ma, Yulong Lu, Jianfeng Lu 외

Training deep neural networks with stochastic gradient descent (SGD) can often achieve zero training loss on real-world tasks although the optimization landscape is known to be highly non-convex. To understand the succes…

A Mean-field Analysis of Deep ResNet and Beyond:Towards Provable Optimization Via Overparameterization From Depth

2020-02-26 · ICLR Workshop DeepDiffEq 2019 12 · Yiping Lu, Chao Ma, Yulong Lu, Jianfeng Lu 외

Training deep neural networks with stochastic gradient descent (SGD) can often achieve zero training loss on real-world tasks although the optimization landscape is known to be highly non-convex. To understand the succes…