paper-with-me

홈 › Papers

On Fast Dropout and its Applicability to Recurrent Networks

2013-11-04 · Justin Bayer, Christian Osendorfer, Daniela Korhammer, Nutan Chen, Sebastian Urban, Patrick van der Smagt

Recurrent Neural Networks (RNNs) are rich models for the processing of sequential data. Recent work on advancing the state of the art has been focused on the optimization or modelling of RNNs, mostly motivated by adressing the problems of the vanishing and exploding gradients. The control of overfitting has seen considerably less attention. This paper contributes to that by analyzing fast dropout, a recent regularization method for generalized linear models and neural networks from a back-propagation inspired perspective. We show that fast dropout implements a quadratic form of an adaptive, per-parameter regularizer, which rewards large weights in the light of underfitting, penalizes them for overconfident predictions and vanishes at minima of an unregularized training loss. The derivatives of that regularizer are exclusively based on the training error signal. One consequence of this is the absense of a global weight attractor, which is particularly appealing for RNNs, since the dynamics are not biased towards a certain regime. We positively test the hypothesis that this improves the performance of RNNs on four musical data sets.

📄 PDF Abstract BibTeX arXiv:1311.0701

Code (1)

George091/RNN tf

Similar Papers 제목 키워드 기반

An ETF view of Dropout regularization

2018-10-14 · Dor Bank, Raja Giryes

Dropout is a popular regularization technique in deep learning. Yet, the reason for its success is still not fully understood. This paper provides a new interpretation of Dropout from a frame theory perspective. By drawi…

Dropout improves Recurrent Neural Networks for Handwriting Recognition

2013-11-05 · Vu Pham, Théodore Bluche, Christopher Kermorvant, Jérôme Louradour

Recurrent neural networks (RNNs) with Long Short-Term memory cells currently hold the best known results in unconstrained handwriting recognition. We show that their performance can be greatly improved using dropout - a …

Handwriting Recognition

Recurrent Dropout without Memory Loss

2016-03-16 · COLING 2016 12 · Stanislau Semeniuta, Aliaksei Severyn, Erhardt Barth

This paper presents a novel approach to recurrent neural network (RNN) regularization. Differently from the widely adopted dropout method, which is applied to \textit{forward} connections of feed-forward architectures or…

A Theoretically Grounded Application of Dropout in Recurrent Neural Networks

2015-12-16 · NeurIPS 2016 12 · Yarin Gal, Zoubin Ghahramani

Recurrent neural networks (RNNs) stand at the forefront of many recent developments in deep learning. Yet a major difficulty with these models is their tendency to overfit, with dropout shown to fail when applied to recu…

Bayesian InferenceDeep LearningLanguage ModellingSentiment Analysis+1

Regularizing Recurrent Networks - On Injected Noise and Norm-based Methods

2014-10-21 · Saahil Ognawala, Justin Bayer

Advancements in parallel processing have lead to a surge in multilayer perceptrons' (MLP) applications and deep learning in the past decades. Recurrent Neural Networks (RNNs) give additional representational power to fee…