paper-with-me

Papers

LightRNN: Memory and Computation-Efficient Recurrent Neural Networks

2016-10-31 · NeurIPS 2016 12 · Xiang Li, Tao Qin, Jian Yang, Tie-Yan Liu

Recurrent neural networks (RNNs) have achieved state-of-the-art performances in many natural language processing tasks, such as language modeling and machine translation. However, when the vocabulary is large, the RNN model will become very big (e.g., possibly beyond the memory capacity of a GPU device) and its training will become very inefficient. In this work, we propose a novel technique to tackle this challenge. The key idea is to use 2-Component (2C) shared embedding for word representations. We allocate every word in the vocabulary into a table, each row of which is associated with a vector, and each column associated with another vector. Depending on its position in the table, a word is jointly represented by two components: a row vector and a column vector. Since the words in the same row share the row vector and the words in the same column share the column vector, we only need $2 \sqrt{|V|}$ vectors to represent a vocabulary of $|V|$ unique words, which are far less than the $|V|$ vectors required by existing approaches. Based on the 2-Component shared embedding, we design a new RNN algorithm and evaluate it using the language modeling task on several benchmark datasets. The results show that our algorithm significantly reduces the model size and speeds up the training process, without sacrifice of accuracy (it achieves similar, if not better, perplexity as compared to state-of-the-art language models). Remarkably, on the One-Billion-Word benchmark Dataset, our algorithm achieves comparable perplexity to previous language models, whilst reducing the model size by a factor of 40-100, and speeding up the training process by a factor of 2. We name our proposed algorithm \emph{LightRNN} to reflect its very small model size and very high training speed.

📄 PDF Abstract BibTeX arXiv:1610.09893

Code (0)

등록된 구현이 없습니다.

Tasks

GPULanguage ModelingLanguage ModellingMachine Translation

Similar Papers 제목 키워드 기반

Fast and Simple Mixture of Softmaxes with BPE and Hybrid-LightRNN for Language Generation

2018-09-25 · Xiang Kong, Qizhe Xie, Zihang Dai, Eduard Hovy

Mixture of Softmaxes (MoS) has been shown to be effective at addressing the expressiveness limitation of Softmax-based models. Despite the known advantage, MoS is practically sealed by its large consumption of memory and…

Image CaptioningMachine TranslationText GenerationTranslation

Optimal Gradient Checkpointing for Sparse and Recurrent Architectures using Off-Chip Memory

2024-12-16 · Wadjih Bencheikh, Jan Finkbeiner, Emre Neftci

Recurrent neural networks (RNNs) are valued for their computational efficiency and reduced memory requirements on tasks involving long sequence lengths but require high memory-processor bandwidth to train. Checkpointing …

Computational Efficiency

Brain-like combination of feedforward and recurrent network components achieves prototype extraction and robust pattern recognition

2022-06-30 · Naresh Balaji Ravichandran, Anders Lansner, Pawel Herman

Associative memory has been a prominent candidate for the computation performed by the massively recurrent neocortical networks. Attractor networks implementing associative memory have offered mechanistic explanation for…

Efficient Real Time Recurrent Learning through combined activity and parameter sparsity

2023-03-10 · Anand Subramoney

Backpropagation through time (BPTT) is the standard algorithm for training recurrent neural networks (RNNs), which requires separate simulation phases for the forward and backward passes for inference and learning, respe…

Simplified Long Short-term Memory Recurrent Neural Networks: part III

2017-07-14 · Atra Akandeh, Fathi M. Salem

This is part III of three-part work. In parts I and II, we have presented eight variants for simplified Long Short Term Memory (LSTM) recurrent neural networks (RNNs). It is noted that fast computation, specially in cons…