paper-with-me

홈 › Papers

The interplay between randomness and structure during learning in RNNs

2020-06-19 · NeurIPS 2020 12 · Friedrich Schuessler, Francesca Mastrogiuseppe, Alexis Dubreuil, Srdjan Ostojic, Omri Barak

Recurrent neural networks (RNNs) trained on low-dimensional tasks have been widely used to model functional biological networks. However, the solutions found by learning and the effect of initial connectivity are not well understood. Here, we examine RNNs trained using gradient descent on different tasks inspired by the neuroscience literature. We find that the changes in recurrent connectivity can be described by low-rank matrices, despite the unconstrained nature of the learning algorithm. To identify the origin of the low-rank structure, we turn to an analytically tractable setting: training a linear RNN on a simplified task. We show how the low-dimensional task structure leads to low-rank changes to connectivity. This low-rank structure allows us to explain and quantify the phenomenon of accelerated learning in the presence of random initial connectivity. Altogether, our study opens a new perspective to understanding trained RNNs in terms of both the learning process and the resulting network structure.

📄 PDF Abstract BibTeX arXiv:2006.11036

Code (1)

frschu/neurips_2020_interplay_randomness_structure 공식 구현 pytorch

Similar Papers 제목 키워드 기반

A unified theory of feature learning in RNNs and DNNs

2026-02-17 · Jan P. Bauer, Kirsten Fischer, Moritz Helias, Agostina Palmigiano arxiv

Recurrent and deep neural networks (RNNs/DNNs) are cornerstone architectures in machine learning. Remarkably, RNNs differ from DNNs only by weight sharing, as can be shown through unrolling in time. How does this structu…

Bayesian Inference

A quantum tug of war between randomness and symmetries on homogeneous spaces

2023-09-11 · Rahul Arvind, Kishor Bharti, Jun Yong Khoo, Dax Enshan Koh 외

We explore the interplay between symmetry and randomness in quantum information. Adopting a geometric approach, we consider states as $H$-equivalent if related by a symmetry transformation characterized by the group $H$.…

Quantum Machine Learning

Domain Compression and its Application to Randomness-Optimal Distributed Goodness-of-Fit

2019-07-20 · Jayadev Acharya, Clément L. Canonne, Yanjun Han, Ziteng Sun 외

We study goodness-of-fit of discrete distributions in the distributed setting, where samples are divided between multiple users who can only release a limited amount of information about their samples due to various info…

Bridging HMMs and RNNs through Architectural Transformations

2018-10-22 · NIPS Workshop IRASL 2018 · Anonymous

A distinct commonality between HMMs and RNNs is that they both learn hidden representations for sequential data. In addition, it has been noted that the backward computation of the Baum-Welch algorithm for HMMs is a spec…

Language ModelingLanguage Modelling

Structure and randomness in planning and reinforcement learning

2021-01-01 · NeurIPS Workshop LMCA 2020 12 · Piotr Kozakowski, Piotr Januszewski, Konrad Czechowski, Łukasz Kuciński 외

Planning in large state spaces inevitably needs to balance depth and breadth of the search. It has a crucial impact on planners performance and most manage this interplay implicitly. We present a novel method $\textit{Sh…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)STS