paper-with-me

Papers

The Nonlinearity Coefficient - Predicting Generalization in Deep Neural Networks

2018-06-01 · ICLR 2019 5 · George Philipp, Jaime G. Carbonell

For a long time, designing neural architectures that exhibit high performance was considered a dark art that required expert hand-tuning. One of the few well-known guidelines for architecture design is the avoidance of exploding gradients, though even this guideline has remained relatively vague and circumstantial. We introduce the nonlinearity coefficient (NLC), a measurement of the complexity of the function computed by a neural network that is based on the magnitude of the gradient. Via an extensive empirical study, we show that the NLC is a powerful predictor of test error and that attaining a right-sized NLC is essential for optimal performance. The NLC exhibits a range of intriguing and important properties. It is closely tied to the amount of information gained from computing a single network gradient. It is tied to the error incurred when replacing the nonlinearity operations in the network with linear operations. It is not susceptible to the confounders of multiplicative scaling, additive bias and layer width. It is stable from layer to layer. Hence, we argue that the NLC is the first robust predictor of overfitting in deep networks.

📄 PDF Abstract BibTeX arXiv:1806.00179

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Stock Price Correlation Coefficient Prediction with ARIMA-LSTM Hybrid Model

2018-08-05 · Hyeong Kyu Choi

Predicting the price correlation of two assets for future time periods is important in portfolio optimization. We apply LSTM recurrent neural networks (RNN) in predicting the stock price correlation coefficient of two in…

Portfolio OptimizationStock Market Prediction

power-law nonlinearity with maximally uniform distribution criterion for improved neural network training in automatic speech recognition

2019-12-22 · Chanwoo Kim, Mehul Kumar, Kwangyoun Kim, Dhananjaya Gowda

In this paper, we describe the Maximum Uniformity of Distribution (MUD) algorithm with the power-law nonlinearity. In this approach, we hypothesize that neural network training will become more stable if feature distribu…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Second-Order Nonlinearity Estimated and Compensated Diffusion LMS Algorithm: Theoretical Upper Bound, Cramer-Rao Lower bound, and Convergence Analysis

2024-03-17 · Hadi Zayyani, Mehdi Korki

In this paper, an algorithm for estimation and compensation of second-order nonlinearity in wireless sensor setwork (WSN) in distributed estimation framework is proposed. First, the effect of second-order nonlinearity on…

Solid Harmonic Wavelet Scattering: Predicting Quantum Molecular Energy from Invariant Descriptors of 3D Electronic Densities

2017-12-01 · NeurIPS 2017 12 · Michael Eickenberg, Georgios Exarchakis, Matthew Hirn, Stephane Mallat

We introduce a solid harmonic wavelet scattering representation, invariant to rigid motion and stable to deformations, for regression and classification of 2D and 3D signals. Solid harmonic wavelets are computed by mul…

General Classificationregression

A new kernel-based approach for overparameterized Hammerstein system identification

2015-04-30 · Riccardo Sven Risuleo, Giulio Bottegal, Håkan Hjalmarsson

In this paper we propose a new identification scheme for Hammerstein systems, which are dynamic systems consisting of a static nonlinearity and a linear time-invariant dynamic system in cascade. We assume that the nonlin…