paper-with-me

홈 › Papers

Phase Diagram of Initial Condensation for Two-layer Neural Networks

2023-03-12 · Zhengan Chen, Yuqing Li, Tao Luo, Zhangchen Zhou, Zhi-Qin John Xu

The phenomenon of distinct behaviors exhibited by neural networks under varying scales of initialization remains an enigma in deep learning research. In this paper, based on the earlier work by Luo et al.~\cite{luo2021phase}, we present a phase diagram of initial condensation for two-layer neural networks. Condensation is a phenomenon wherein the weight vectors of neural networks concentrate on isolated orientations during the training process, and it is a feature in non-linear learning process that enables neural networks to possess better generalization abilities. Our phase diagram serves to provide a comprehensive understanding of the dynamical regimes of neural networks and their dependence on the choice of hyperparameters related to initialization. Furthermore, we demonstrate in detail the underlying mechanisms by which small initialization leads to condensation at the initial training stage.

📄 PDF Abstract BibTeX arXiv:2303.06561

Code (0)

등록된 구현이 없습니다.

Tasks

Vocal Bursts Valence Prediction

Similar Papers 제목 키워드 기반

Empirical Phase Diagram for Three-layer Neural Networks with Infinite Width

2022-05-24 · Hanxu Zhou, Qixuan Zhou, Zhenyuan Jin, Tao Luo 외

Substantial work indicates that the dynamics of neural networks (NNs) is closely related to their initialization of parameters. Inspired by the phase diagram for two-layer ReLU NNs with infinite width (Luo et al., 2021),…

On the dynamics of three-layer neural networks: initial condensation

2024-02-25 · Zheng-an Chen, Tao Luo

Empirical and theoretical works show that the input weights of two-layer neural networks, when initialized with small values, converge towards isolated orientations. This phenomenon, referred to as condensation, indicate…

Phase diagram for two-layer ReLU neural networks at infinite-width limit

2020-07-15 · Tao Luo, Zhi-Qin John Xu, Zheng Ma, Yaoyu Zhang

How neural network behaves during the training over different choices of hyperparameters is an important question in the study of neural networks. In this work, inspired by the phase diagram in statistical mechanics, we …

Towards Understanding the Condensation of Neural Networks at Initial Training

2021-05-25 · Hanxu Zhou, Qixuan Zhou, Tao Luo, Yaoyu Zhang 외

Empirical works show that for ReLU neural networks (NNs) with small initialization, input weights of hidden neurons (the input weight of a hidden neuron consists of the weight from its input layer to the hidden neuron an…

Understanding the Initial Condensation of Convolutional Neural Networks

2023-05-17 · Zhangchen Zhou, Hanxu Zhou, Yuqing Li, Zhi-Qin John Xu

Previous research has shown that fully-connected networks with small initialization and gradient-based training methods exhibit a phenomenon known as condensation during training. This phenomenon refers to the input weig…