paper-with-me

홈 › Papers

Beyond Random Matrix Theory for Deep Networks

2020-06-13 · Diego Granziol

We investigate whether the Wigner semi-circle and Marcenko-Pastur distributions, often used for deep neural network theoretical analysis, match empirically observed spectral densities. We find that even allowing for outliers, the observed spectral shapes strongly deviate from such theoretical predictions. This raises major questions about the usefulness of these models in deep learning. We further show that theoretical results, such as the layered nature of critical points, are strongly dependent on the use of the exact form of these limiting spectral densities. We consider two new classes of matrix ensembles; random Wigner/Wishart ensemble products and percolated Wigner/Wishart ensembles, both of which better match observed spectra. They also give large discrete spectral peaks at the origin, providing a theoretical explanation for the observation that various optima can be connected by one dimensional of low loss values. We further show that, in the case of a random matrix product, the weight of the discrete spectral component at $0$ depends on the ratio of the dimensions of the weight matrices.

📄 PDF Abstract BibTeX arXiv:2006.07721

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Singular Value Decomposition, Applications and Beyond

2015-10-29 · Zhihua Zhang

The singular value decomposition (SVD) is not only a classical theory in matrix computation and analysis, but also is a powerful tool in machine learning and modern data analysis. In this tutorial we first study the basi…

BIG-bench Machine LearningMatrix Completion

On the phase diagram of extensive-rank symmetric matrix denoising beyond rotational invariance

2024-11-04 · Jean Barbier, Francesco Camilli, Justin Ko, Koki Okajima

Matrix denoising is central to signal processing and machine learning. Its statistical analysis when the matrix to infer has a factorised structure with a rank growing proportionally to its dimension remains a challenge,…

Denoising

Random matrix theory and the loss surfaces of neural networks

2023-06-03 · Nicholas P Baskerville

Neural network models are one of the most successful approaches to machine learning, enjoying an enormous amount of development and research over recent years and finding concrete real-world applications in almost any co…

Nonlinear random matrix theory for deep learning

2017-12-01 · NeurIPS 2017 12 · Jeffrey Pennington, Pratik Worah

Neural network configurations with random weights play an important role in the analysis of deep learning. They define the initial loss landscape and are closely related to kernel and random feature methods. Despite the …

Deep LearningMemorization

A Random Matrix Theory Perspective on the Spectrum of Learned Features and Asymptotic Generalization Capabilities

2024-10-24 · Yatin Dandi, Luca Pesce, Hugo Cui, Florent Krzakala 외

A key property of neural networks is their capacity of adapting to data during training. Yet, our current mathematical understanding of feature learning and its relationship to generalization remain limited. In this work…