paper-with-me

홈 › Papers

Implicit recurrent networks: A novel approach to stationary input processing with recurrent neural networks in deep learning

2020-10-20 · Sebastian Sanokowski

The brain cortex, which processes visual, auditory and sensory data in the brain, is known to have many recurrent connections within its layers and from higher to lower layers. But, in the case of machine learning with neural networks, it is generally assumed that strict feed-forward architectures are suitable for static input data, such as images, whereas recurrent networks are required mainly for the processing of sequential input, such as language. However, it is not clear whether also processing of static input data benefits from recurrent connectivity. In this work, we introduce and test a novel implementation of recurrent neural networks with lateral and feed-back connections into deep learning. This departure from the strict feed-forward structure prevents the use of the standard error backpropagation algorithm for training the networks. Therefore we provide an algorithm which implements the backpropagation algorithm on a implicit implementation of recurrent networks, which is different from state-of-the-art implementations of recurrent neural networks. Our method, in contrast to current recurrent neural networks, eliminates the use of long chains of derivatives due to many iterative update steps, which makes learning computationally less costly. It turns out that the presence of recurrent intra-layer connections within a one-layer implicit recurrent network enhances the performance of neural networks considerably: A single-layer implicit recurrent network is able to solve the XOR problem, while a feed-forward network with monotonically increasing activation function fails at this task. Finally, we demonstrate that a two-layer implicit recurrent architecture leads to a better performance in a regression task of physical parameters from the measured trajectory of a damped pendulum.

📄 PDF Abstract BibTeX arXiv:2010.10564

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Memory and forecasting capacities of nonlinear recurrent networks

2020-04-22 · Lukas Gonon, Lyudmila Grigoryeva, Juan-Pablo Ortega

The notion of memory capacity, originally introduced for echo state and linear networks with independent inputs, is generalized to nonlinear recurrent networks with stationary but dependent inputs. The presence of depend…

Time SeriesTime Series Analysis

The Importance of Being Recurrent for Modeling Hierarchical Structure

2018-03-09 · EMNLP 2018 10 · Ke Tran, Arianna Bisazza, Christof Monz

Recent work has shown that recurrent neural networks (RNNs) can implicitly capture and exploit hierarchical information when trained to solve common natural language processing tasks such as language modeling (Linzen et …

Language ModelingLanguage ModellingMachine TranslationTranslation

Statistical and Neural Network Based Speech Activity Detection in Non-Stationary Acoustic Environments

2020-07-28

Speech activity detection (SAD), which often rests on the fact that the noise is "more" stationary than speech, is particularly challenging in non-stationary environments, because the time variance of the acoustic scene …

Action DetectionActivity Detection

Interneurons accelerate learning dynamics in recurrent neural networks for statistical adaptation

2022-09-21 · David Lipshutz, Cengiz Pehlevan, Dmitri B. Chklovskii

Early sensory systems in the brain rapidly adapt to fluctuating input statistics, which requires recurrent communication between neurons. Mechanistically, such recurrent communication is often indirect and mediated by lo…

Input correlations impede suppression of chaos and learning in balanced rate networks

2022-01-24 · Rainer Engelken, Alessandro Ingrosso, Ramin Khajeh, Sven Goedeke 외

Neural circuits exhibit complex activity patterns, both spontaneously and evoked by external stimuli. Information encoding and learning in neural circuits depend on how well time-varying stimuli can control spontaneous n…