paper-with-me

홈 › Papers

On the Impact of Stable Ranks in Deep Nets

2021-10-05 · Bogdan Georgiev, Lukas Franken, Mayukh Mukherjee, Georgios Arvanitidis

A recent line of work has established intriguing connections between the generalization/compression properties of a deep neural network (DNN) model and the so-called layer weights' stable ranks. Intuitively, the latter are indicators of the effective number of parameters in the net. In this work, we address some natural questions regarding the space of DNNs conditioned on the layers' stable rank, where we study feed-forward dynamics, initialization, training and expressivity. To this end, we first propose a random DNN model with a new sampling scheme based on stable rank. Then, we show how feed-forward maps are affected by the constraint and how training evolves in the overparametrized regime (via Neural Tangent Kernels). Our results imply that stable ranks appear layerwise essentially as linear factors whose effect accumulates exponentially depthwise. Moreover, we provide empirical analysis suggesting that stable rank initialization alone can lead to convergence speed ups.

📄 PDF Abstract BibTeX arXiv:2110.02333

Code (0)

등록된 구현이 없습니다.

Tasks

Natural Questions

Similar Papers 제목 키워드 기반

Metrics for Learning in Topological Persistence

2019-06-11 · Henri Riihimäki, José Licón-Saláiz

Persistent homology analysis provides means to capture the connectivity structure of data sets in various dimensions. On the mathematical level, by defining a metric between the objects that persistence attaches to data …

General Classification

Testing Zipf’s meaning-frequency law with wordnets as sense inventories

2019-07-01 · GWC 2019 7 · Francis Bond, Arkadiusz Janz, Marek Maziarz, Ewa Rudnicka

According to George K. Zipf, more frequent words have more senses. We have tested this law using corpora and wordnets of English, Spanish, Portuguese, French, Polish, Japanese, Indonesian and Chinese. We have proved that…

LEMMA

Singular Value Perturbation and Deep Network Optimization

2022-03-07 · Rudolf H. Riedi, Randall Balestriero, Richard G. Baraniuk

We develop new theoretical results on matrix perturbation to shed light on the impact of architecture on the performance of a deep network. In particular, we explain analytically what deep learning practitioners have lon…

The impact of feature importance methods on the interpretation of defect classifiers

2022-02-04 · Gopi Krishnan Rajbahadur, Shaowei Wang, Yasutaka Kamei, Ahmed E. Hassan

Classifier specific (CS) and classifier agnostic (CA) feature importance methods are widely used (often interchangeably) by prior studies to derive feature importance ranks from a defect classifier. However, different fe…

Feature Importance

Improve Ranking Correlation of Super-net through Training Scheme from One-shot NAS to Few-shot NAS

2022-06-13 · Jiawei Liu, Kaiyu Zhang, Weitai Hu, Qing Yang

The algorithms of one-shot neural architecture search(NAS) have been widely used to reduce computation consumption. However, because of the interference among the subnets in which weights are shared, the subnets inherite…

Neural Architecture Search