paper-with-me

홈 › Papers

Enhancing Accuracy in Deep Learning Using Random Matrix Theory

2023-10-04 · Leonid Berlyand, Etienne Sandier, Yitzchak Shmalo, Lei Zhang

We explore the applications of random matrix theory (RMT) in the training of deep neural networks (DNNs), focusing on layer pruning that is reducing the number of DNN parameters (weights). Our numerical results show that this pruning leads to a drastic reduction of parameters while not reducing the accuracy of DNNs and CNNs. Moreover, pruning the fully connected DNNs actually increases the accuracy and decreases the variance for random initializations. Our numerics indicate that this enhancement in accuracy is due to the simplification of the loss landscape. We next provide rigorous mathematical underpinning of these numerical results by proving the RMT-based Pruning Theorem. Our results offer valuable insights into the practical application of RMT for the creation of more efficient and accurate deep-learning models.

📄 PDF Abstract BibTeX arXiv:2310.03165

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Learning

Methods 이 논문이 사용한 방법론

Pruning 설명 없음

Similar Papers 제목 키워드 기반

Using Empirical Covariance Matrix in Enhancing Prediction Accuracy of Linear Models with Missing Information

2016-11-21 · Ahmadreza Moradipari, Sina Shahsavari, Ashkan Esmaeili, Farokh Marvasti

Inference and Estimation in Missing Information (MI) scenarios are important topics in Statistical Learning Theory and Machine Learning (ML). In ML literature, attempts have been made to enhance prediction through precis…

feature selectionLearning TheoryMatrix Completion

Random matrix theory and the loss surfaces of neural networks

2023-06-03 · Nicholas P Baskerville

Neural network models are one of the most successful approaches to machine learning, enjoying an enormous amount of development and research over recent years and finding concrete real-world applications in almost any co…

Nonlinear random matrix theory for deep learning

2017-12-01 · NeurIPS 2017 12 · Jeffrey Pennington, Pratik Worah

Neural network configurations with random weights play an important role in the analysis of deep learning. They define the initial loss landscape and are closely related to kernel and random feature methods. Despite the …

Deep LearningMemorization

Detecting overfitting in Neural Networks during long-horizon grokking using Random Matrix Theory

2026-05-12 · Hari K. Prakash, Charles H Martin arxiv

Training Neural Networks (NNs) without overfitting is difficult; detecting that overfitting is difficult as well. We present a novel Random Matrix Theory method that detects the onset of overfitting in deep learning mode…

Matricial Free Energy as a Gaussianizing Regularizer: Enhancing Autoencoders for Gaussian Code Generation

2025-10-20 · Rishi Sonthalia, Raj Rao Nadakuditi arxiv

We introduce a novel regularization scheme for autoencoders based on matricial free energy. Our approach defines a differentiable loss function in terms of the singular values of the code matrix (code dimension x batch s…

Code Generation