paper-with-me

홈 › Papers

CERSA: Cumulative Energy-Retaining Subspace Adaptation for Memory-Efficient Fine-Tuning

2026-05-05 · Jingze Ge, Xue Geng, Yun Liu, Wanqi Dong, Wang Zhe Mark, Min Wu, Ngai-Man Cheung, Bharadwaj Veeravalli, Xulei Yang arxiv

To mitigate the memory constraints associated with fine-tuning large pre-trained models, existing parameter-efficient fine-tuning (PEFT) methods, such as LoRA, rely on low-rank updates. However, such updates fail to fully capture the rank characteristics of the weight modifications observed in full-parameter fine-tuning, resulting in a performance gap. Furthermore, LoRA and other existing PEFT methods still require substantial memory to store the full set of frozen weights, limiting their efficiency in resource-constrained settings. To addres these limitations, we introduce Cumulative Energy-Retaining Subspace Adaptation (CERSA), a novel fine-tuning paradigm that leverages singular value decomposition (SVD) to retain only the principal components responsible for 90% to 95% of the spectral energy. By fine-tuning low-rank representations derived from this principal subspace, CERSA significantly reduces memory consumption. We conduct extensive evaluations of CERSA across models of varying scales and domains, including image recognition, text-to-image generation, and natural language understanding. Empirical results demonstrate that CERSA consistently outperforms state-of-the-art PEFT methods while achieving substantially lower memory requirements. The code will be publicly released.

📄 PDF Abstract BibTeX arXiv:2605.08174

Code (0)

등록된 구현이 없습니다.

Tasks

parameter-efficient fine-tuningNatural Language UnderstandingText-to-Image Generation

Similar Papers 제목 키워드 기반

Multi-step Online Unsupervised Domain Adaptation

2020-02-20 · J. H. Moon, Debasmit Das, C. S. George Lee

In this paper, we address the Online Unsupervised Domain Adaptation (OUDA) problem, where the target data are unlabelled and arriving sequentially. The traditional methods on the OUDA problem mainly focus on transforming…

Domain AdaptationOnline unsupervised domain adaptationUnsupervised Domain Adaptation

SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models

2026-04-20 · Dongxin Guo, Jikun Wu, Siu Ming Yiu arxiv

Safety alignment in large language models is remarkably shallow: it is concentrated in the first few output tokens and reversible by fine-tuning on as few as 100 adversarial examples. This fragility becomes critical in r…

Domain Adaptation

Subspace Node Pruning

2024-05-26 · Joshua Offergeld, Marcel van Gerven, Nasir Ahmad

Efficiency of neural network inference is undeniably important in a time where commercial use of AI models increases daily. Node pruning is the art of removing computational units such as neurons, filters, attention head…

Distribution-Specific Curvature Control with Finite-Sample Guarantees for Open-Weight Safety

2026-07-24 · Domenic Rosati, Ali Dadsetan, Hong Huang, Xijie Zeng 외 arxiv

A short fine-tuning run can undo the safety guards of an open-weight model---retraining a refusal-trained assistant to aid weapons development or produce hate speech. Preventing such harmful fine-tuning while retaining b…

Dual SVM Training on a Budget

2018-06-26 · Sahar Qaadan, Merlin Schüler, Tobias Glasmachers

We present a dual subspace ascent algorithm for support vector machine training that respects a budget constraint limiting the number of support vectors. Budget methods are effective for reducing the training time of ker…