paper-with-me

홈 › Papers

Layer Flexible Adaptive Computational Time

2018-12-06 · Lida Zhang, Abdolghani Ebrahimi, Diego Klabjan

Deep recurrent neural networks perform well on sequence data and are the model of choice. However, it is a daunting task to decide the structure of the networks, i.e. the number of layers, especially considering different computational needs of a sequence. We propose a layer flexible recurrent neural network with adaptive computation time, and expand it to a sequence to sequence model. Different from the adaptive computation time model, our model has a dynamic number of transmission states which vary by step and sequence. We evaluate the model on a financial data set and Wikipedia language modeling. Experimental results show the performance improvement of 7\% to 12\% and indicate the model's ability to dynamically change the number of layers along with the computational steps.

📄 PDF Abstract BibTeX arXiv:1812.02335

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage Modelling

Similar Papers 제목 키워드 기반

Layer Flexible Adaptive Computation Time for Recurrent Neural Networks

2019-09-25 · Lida Zhang, Diego Klabjan

Deep recurrent neural networks perform well on sequence data and are the model of choice. However, it is a daunting task to decide the structure of the networks, i.e. the number of layers, especially considering differen…

Language ModelingLanguage Modelling

Flexible Mixed Precision Quantization for Learned Image Compression

2025-06-02 · Md Adnan Faisal Hossain, Zhihao Duan, Fengqing Zhu

Despite its improvements in coding performance compared to traditional codecs, Learned Image Compression (LIC) suffers from large computational costs for storage and deployment. Model quantization offers an effective sol…

Image CompressionQuantization

F\textsuperscript{2}LP-AP: Fast \& Flexible Label Propagation with Adaptive Propagation Kernel

2026-04-22 · Yutong Shen, Ruizhe Xia, Jingyi Liu, Yinqi Liu arxiv

Semi-supervised node classification is a foundational task in graph machine learning, yet state-of-the-art Graph Neural Networks (GNNs) are hindered by significant computational overhead and reliance on strong homophily …

Computational EfficiencyNode Classification

DMSC: Dynamic Multi-Scale Coordination Framework for Time Series Forecasting

2025-08-03 · Haonan Yang, Jianchao Tang, Zhuo Li, Long Lan arxiv

Time Series Forecasting (TSF) faces persistent challenges in modeling intricate temporal dependencies across different scales. Despite recent advances leveraging different decomposition operations and novel architectures…

Computational EfficiencyTime Series Forecasting

QTALE: Quantization-Robust Token-Adaptive Layer Execution for LLMs

2026-02-11 · Kanghyun Noh, Jinheon Choi, Yulhwa Kim arxiv

Large language models (LLMs) demand substantial computational and memory resources, posing challenges for efficient deployment. Two complementary approaches have emerged to address these issues: token-adaptive layer exec…