paper-with-me

Papers

Deep Learning with Eigenvalue Decay Regularizer

2016-04-24 · Oswaldo Ludwig

This paper extends our previous work on regularization of neural networks using Eigenvalue Decay by employing a soft approximation of the dominant eigenvalue in order to enable the calculation of its derivatives in relation to the synaptic weights, and therefore the application of back-propagation, which is a primary demand for deep learning. Moreover, we extend our previous theoretical analysis to deep neural networks and multiclass classification problems. Our method is implemented as an additional regularizer in Keras, a modular neural networks library written in Python, and evaluated in the benchmark data sets Reuters Newswire Topics Classification, IMDB database for binary sentiment classification, MNIST database of handwritten digits and CIFAR-10 data set for image classification.

📄 PDF Abstract BibTeX arXiv:1604.06985

Code (1)

oswaldoludwig/Eigenvalue-Decay-Regularizer-for-Keras 공식 구현 tf

Tasks

ClassificationDeep LearningGeneral Classificationimage-classificationImage ClassificationSentiment AnalysisSentiment Classification

Similar Papers 제목 키워드 기반

Optimal rates for the regularized learning algorithms under general source condition

2016-11-07 · Abhishake Rastogi, Sivananthan Sampath

We consider the learning algorithms under general source condition with the polynomial decay of the eigenvalues of the integral operator in vector-valued function setting. We discuss the upper convergence rates of Tikhon…

Evolution of Eigenvalue Decay in Deep Networks

2019-05-28 · Lukas Pfahler, Katharina Morik

The linear transformations in converged deep networks show fast eigenvalue decay. The distribution of eigenvalues looks like a Heavy-tail distribution, where the vast majority of eigenvalues is small, but not actually ze…

Eigenvalue Decay Implies Polynomial-Time Learnability for Neural Networks

2017-08-11 · NeurIPS 2017 12 · Surbhi Goel, Adam Klivans

We consider the problem of learning function classes computed by neural networks with various activations (e.g. ReLU or Sigmoid), a task believed to be computationally intractable in the worst-case. A major open problem …

Relaxed Sparse Eigenvalue Conditions for Sparse Estimation via Non-convex Regularized Regression

2013-06-14 · Zheng Pan, Chang-Shui Zhang

Non-convex regularizers usually improve the performance of sparse estimation in practice. To prove this fact, we study the conditions of sparse estimations for the sharp concave regularizers which are a general family of…

parameter estimationregression

On the Benefits of Active Data Collection in Operator Learning

2024-10-25 · Unique Subedi, Ambuj Tewari

We investigate active data collection strategies for operator learning when the target operator is linear and the input functions are drawn from a mean-zero stochastic process with continuous covariance kernels. With an …

Operator learning