paper-with-me

홈 › Papers

Delta Tuning: A Comprehensive Study of Parameter Efficient Methods for Pre-trained Language Models

2022-03-14 · Ning Ding, Yujia Qin, Guang Yang, Fuchao Wei, Zonghan Yang, Yusheng Su, Shengding Hu, Yulin Chen, Chi-Min Chan, Weize Chen, Jing Yi, Weilin Zhao, Xiaozhi Wang, Zhiyuan Liu, Hai-Tao Zheng, Jianfei Chen, Yang Liu, Jie Tang, Juanzi Li, Maosong Sun

Despite the success, the process of fine-tuning large-scale PLMs brings prohibitive adaptation costs. In fact, fine-tuning all the parameters of a colossal model and retaining separate instances for different tasks are practically infeasible. This necessitates a new branch of research focusing on the parameter-efficient adaptation of PLMs, dubbed as delta tuning in this paper. In contrast with the standard fine-tuning, delta tuning only fine-tunes a small portion of the model parameters while keeping the rest untouched, largely reducing both the computation and storage costs. Recent studies have demonstrated that a series of delta tuning methods with distinct tuned parameter selection could achieve performance on a par with full-parameter fine-tuning, suggesting a new promising way of stimulating large-scale PLMs. In this paper, we first formally describe the problem of delta tuning and then comprehensively review recent delta tuning approaches. We also propose a unified categorization criterion that divide existing delta tuning methods into three groups: addition-based, specification-based, and reparameterization-based methods. Though initially proposed as an efficient method to steer large models, we believe that some of the fascinating evidence discovered along with delta tuning could help further reveal the mechanisms of PLMs and even deep neural networks. To this end, we discuss the theoretical principles underlying the effectiveness of delta tuning and propose frameworks to interpret delta tuning from the perspective of optimization and optimal control, respectively. Furthermore, we provide a holistic empirical study of representative methods, where results on over 100 NLP tasks demonstrate a comprehensive performance comparison of different approaches. The experimental results also cover the analysis of combinatorial, scaling and transferable properties of delta tuning.

📄 PDF Abstract BibTeX arXiv:2203.06904

Code (1)

thunlp/opendelta 공식 구현 pytorch

Tasks

Text Classification

Similar Papers 제목 키워드 기반

OpenDelta: A Plug-and-play Library for Parameter-efficient Adaptation of Pre-trained Models

2023-07-05 · Shengding Hu, Ning Ding, Weilin Zhao, Xingtai Lv 외

The scale of large pre-trained models (PTMs) poses significant challenges in adapting to downstream tasks due to the high optimization overhead and storage costs associated with full-parameter fine-tuning. To address thi…

DARE the Extreme: Revisiting Delta-Parameter Pruning For Fine-Tuned Models

2024-10-12 · Wenlong Deng, Yize Zhao, Vala Vakilian, Minghui Chen 외

Storing open-source fine-tuned models separately introduces redundancy and increases response times in applications utilizing multiple models. Delta-parameter pruning (DPP), particularly the random drop and rescale (DARE…

CoLAparameter-efficient fine-tuning

Delta-LoRA: Fine-Tuning High-Rank Parameters with the Delta of Low-Rank Matrices

2023-09-05 · Bojia Zi, Xianbiao Qi, Lingzhi Wang, Jianan Wang 외

In this paper, we present Delta-LoRA, which is a novel parameter-efficient approach to fine-tune large language models (LLMs). In contrast to LoRA and other low-rank adaptation methods such as AdaLoRA, Delta-LoRA not onl…

Sparse Structure Search for Delta Tuning

2022-11-01 · NIPS 2022 11 · Shengding Hu, Zhen Zhang, Ning Ding, Yadao Wang 외

Adapting large pre-trained models (PTMs) through fine-tuning imposes prohibitive computational and storage burdens. Recent studies of delta tuning (DT), i.e., parameter-efficient tuning, find that only optimizing a smal…

Safe Delta: Consistently Preserving Safety when Fine-Tuning LLMs on Diverse Datasets

2025-05-17 · Ning Lu, Shengcai Liu, Jiahao Wu, WeiYu Chen 외

Large language models (LLMs) have shown great potential as general-purpose AI assistants across various domains. To fully leverage this potential in specific applications, many companies provide fine-tuning API services,…