paper-with-me

홈 › Papers

Equivalence in Deep Neural Networks via Conjugate Matrix Ensembles

2020-06-14 · Mehmet Süzen

A numerical approach is developed for detecting the equivalence of deep learning architectures. The method is based on generating Mixed Matrix Ensembles (MMEs) out of deep neural network weight matrices and {\it conjugate circular ensemble} matching the neural architecture topology. Following this, the empirical evidence supports the {\it phenomenon} that difference between spectral densities of neural architectures and corresponding {\it conjugate circular ensemble} are vanishing with different decay rates at the long positive tail part of the spectrum i.e., cumulative Circular Spectral Difference (CSD). This finding can be used in establishing equivalences among different neural architectures via analysis of fluctuations in CSD. We investigated this phenomenon for a wide range of deep learning vision architectures and with circular ensembles originating from statistical quantum mechanics. Practical implications of the proposed method for artificial and natural neural architectures discussed such as the possibility of using the approach in Neural Architecture Search (NAS) and classification of biological neural networks.

📄 PDF Abstract BibTeX arXiv:2006.13687

Code (1)

msuzen/bristol pytorch

Tasks

Neural Architecture Search

Methods 이 논문이 사용한 방법론

Sigmoid Activation 설명 없음
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Tanh Activation 설명 없음
LSTM An LSTM is a type of recurrent neural network that addresses the vanishing gradient problem in vanilla…

Similar Papers 제목 키워드 기반

Largest Eigenvalues of the Conjugate Kernel of Single-Layered Neural Networks

2022-01-13 · Lucas Benigni, Sandrine Péché

This paper is concerned with the asymptotic distribution of the largest eigenvalues for some nonlinear random matrix ensemble stemming from the study of neural networks. More precisely we consider $M= \frac{1}{m} YY^\top…

Eigen-Spike Emergence and Quadratic Equivalents for Conjugate Kernels on Nonlinearly Separable Data

2026-05-28 · Collin Cranston, Zhichao Wang, Todd Kemp, Michael W. Mahoney arxiv

Recent work in random matrix theory (RMT) has developed the notion of deterministic equivalents: typically linear surrogate models that approximate the spectral behavior of large nonlinear random matrices, such as nonlin…

Dynamical ensembles equivalence in fluid mechanics

1996-05-09 · Giovanni Gallavotti

Dissipative Euler and Navier Stokes equations are discussed with the aim of proposing several experiments apt to test the equivalence of dynamical ensembles and the chaotic hypothesis.

Kullback-Leibler Proximal Variational Inference

2015-12-01 · NeurIPS 2015 12 · Mohammad E. Khan, Pierre Baque, François Fleuret, Pascal Fua

We propose a new variational inference method based on the Kullback-Leibler (KL) proximal term. We make two contributions towards improving efficiency of variational inference. Firstly, we derive a KL proximal-point algo…

Variational Inference

Riemannian statistics meets random matrix theory: towards learning from high-dimensional covariance matrices

2022-03-01 · Salem Said, Simon Heuveline, Cyrus Mostajeran

Riemannian Gaussian distributions were initially introduced as basic building blocks for learning models which aim to capture the intrinsic structure of statistical populations of positive-definite matrices (here called …