paper-with-me

홈 › Papers

On the Landscape of One-hidden-layer Sparse Networks and Beyond

2020-09-16 · Dachao Lin, Ruoyu Sun, Zhihua Zhang

Sparse neural networks have received increasing interest due to their small size compared to dense networks. Nevertheless, most existing works on neural network theory have focused on dense neural networks, and the understanding of sparse networks is very limited. In this paper, we study the loss landscape of one-hidden-layer sparse networks. First, we consider sparse networks with a dense final layer. We show that linear networks can have no spurious valleys under special sparse structures, and non-linear networks could also admit no spurious valleys under a wide final layer. Second, we discover that spurious valleys and spurious minima can exist for wide sparse networks with a sparse final layer. This is different from wide dense networks which do not have spurious valleys under mild assumptions.

📄 PDF Abstract BibTeX arXiv:2009.07439

Code (0)

등록된 구현이 없습니다.

Tasks

Network Pruning

Similar Papers 제목 키워드 기반

Geometry of Learning -- L2 Phase Transitions in Deep and Shallow Neural Networks

2025-05-10 · Ibrahim Talha Ersoy, Karoline Wiesner

When neural networks (NNs) are subject to L2 regularization, increasing the regularization strength beyond a certain threshold pushes the model into an under-parameterization regime. This transition manifests as a first-…

L2 Regularization

How to Characterize The Landscape of Overparameterized Convolutional Neural Networks

2020-12-01 · NeurIPS 2020 12 · Yihong Gu, Weizhong Zhang, Cong Fang, Jason D. Lee 외

For many initialization schemes, parameters of two randomly initialized deep neural networks (DNNs) can be quite different, but feature distributions of the hidden nodes are similar at each layer. With the help of a new …

Unveiling Hidden Convexity in Deep Learning: a Sparse Signal Processing Perspective

2026-03-25 · Emi Zeger, Mert Pilanci arxiv

Deep neural networks (DNNs), particularly those using Rectified Linear Unit (ReLU) activation functions, have achieved remarkable success across diverse machine learning tasks, including image recognition, audio processi…

On the Landscape of Sparse Linear Networks

2021-01-01 · Dachao Lin, Ruoyu Sun, Zhihua Zhang

Network pruning, or sparse network has a long history and practical significance in modern applications. Although the loss functions of neural networks may yield bad landscape due to non-convexity, we focus on linear act…

Network Pruning

Weight-space symmetry in neural network loss landscapes revisited

2019-09-25 · Berfin Simsek, Johanni Brea, Bernd Illing, Wulfram Gerstner

Neural network training depends on the structure of the underlying loss landscape, i.e. local minima, saddle points, flat plateaus, and loss barriers. In relation to the structure of the landscape, we study the permutati…