paper-with-me

Papers

Variator: Accelerating Pre-trained Models with Plug-and-Play Compression Modules

2023-10-24 · Chaojun Xiao, Yuqi Luo, Wenbin Zhang, Pengle Zhang, Xu Han, Yankai Lin, Zhengyan Zhang, Ruobing Xie, Zhiyuan Liu, Maosong Sun, Jie zhou

Pre-trained language models (PLMs) have achieved remarkable results on NLP tasks but at the expense of huge parameter sizes and the consequent computational costs. In this paper, we propose Variator, a parameter-efficient acceleration method that enhances computational efficiency through plug-and-play compression plugins. Compression plugins are designed to reduce the sequence length via compressing multiple hidden vectors into one and trained with original PLMs frozen. Different from traditional model acceleration methods, which compress PLMs to smaller sizes, Variator offers two distinct advantages: (1) In real-world applications, the plug-and-play nature of our compression plugins enables dynamic selection of different compression plugins with varying acceleration ratios based on the current workload. (2) The compression plugin comprises a few compact neural network layers with minimal parameters, significantly saving storage and memory overhead, particularly in scenarios with a growing number of tasks. We validate the effectiveness of Variator on seven datasets. Experimental results show that Variator can save 53% computational costs using only 0.9% additional parameters with a performance drop of less than 2%. Moreover, when the model scales to billions of parameters, Variator matches the strong performance of uncompressed PLMs.

📄 PDF Abstract BibTeX arXiv:2310.15724

Code (1)

thunlp/compression-plugin 공식 구현 pytorch

Tasks

Computational Efficiency

Similar Papers 제목 키워드 기반

New class of compounds - variators - are reprogramming substrate specificity of H4K12Ac, H4K16Ac and H4K20Ac epigenetic marks reading bromodomain of BPTF protein

2015-06-23

Previously reported [http://arxiv.org/abs/1506.06433] reprogramming of substrate specificity of H3K4Me3 epigenetic marks reading PHD domain of BPTF protein illustrates therapeutic potential of a new class of non-inhibito…

Specificity

DLFR-VAE: Dynamic Latent Frame Rate VAE for Video Generation

2025-02-17 · Zhihang Yuan, Siyuan Wang, Rui Xie, Hanling Zhang 외

In this paper, we propose the Dynamic Latent Frame Rate VAE (DLFR-VAE), a training-free paradigm that can make use of adaptive temporal compression in latent space. While existing video generative models apply fixed comp…

Video Generation

PeRFlow: Piecewise Rectified Flow as Universal Plug-and-Play Accelerator

2024-05-13 · Hanshu Yan, Xingchao Liu, Jiachun Pan, Jun Hao Liew 외

We present Piecewise Rectified Flow (PeRFlow), a flow-based method for accelerating diffusion models. PeRFlow divides the sampling process of generative flows into several time windows and straightens the trajectories in…

LLMC: Benchmarking Large Language Model Quantization with a Versatile Compression Toolkit

2024-05-09 · Ruihao Gong, Yang Yong, Shiqiao Gu, Yushi Huang 외

Recent advancements in large language models (LLMs) are propelling us toward artificial general intelligence with their remarkable emergent abilities and reasoning capabilities. However, the substantial computational and…

BenchmarkingComputational EfficiencyLanguage ModelingLanguage Modelling+2

New class of compounds - variators - are reprogramming substrate specificity of H3K4me3 epigenetic marks reading PHD domain of BPTF protein

2015-06-22

In lymphoma, mutations in genes of histone modifying proteins are frequently observed. Notably, somatic mutations in the activatory histone modification writing protein MLL2 and the repressive modification writer EZH2 ar…

Specificity