paper-with-me

홈 › Papers

Label-Free Cross-Task LoRA Merging with Null-Space Compression

2026-03-27 · Wonyoung Lee, Wooseong Jeong, Kuk-Jin Yoon arxiv

Model merging combines independently fine-tuned checkpoints without joint multi-task training. In the era of foundation-model, fine-tuning with Low-Rank Adaptation (LoRA) is prevalent, making LoRA merging a promising target. Existing approaches can work in homogeneous settings where all target tasks are classification but often fail when tasks span classification and regression. Approaches using entropy-based surrogates do not apply to regression and are costly for large language models due to long token sequences. We introduce Null-Space Compression (NSC) Merging, a label-free, output-agnostic method that sets merge weights from adapter geometry. Our key observation is that during LoRA finetuning the down-projection factor $A$ in $ΔW = BA$ compresses its null space, and the compression correlates with performance. NSC uses this as an optimization signal for merging that can generalize across classification, regression, and sequence generation. NSC achieves state-of-the-art performance across twenty heterogeneous vision tasks with balanced gains where prior methods overfit subsets of tasks. It also outperforms baselines on six NLI benchmarks and on vision-language evaluations for VQA and image captioning, demonstrating scalability and effectiveness.

📄 PDF Abstract BibTeX arXiv:2603.26317

Code (0)

등록된 구현이 없습니다.

Tasks

Image Captioning

Similar Papers 제목 키워드 기반

Decouple and Orthogonalize: A Data-Free Framework for LoRA Merging

2025-05-21 · Shenghe Zheng, Hongzhi Wang, Chenyu Huang, Xiaohui Wang 외

With more open-source models available for diverse tasks, model merging has gained attention by combining models into one, reducing training, storage, and inference costs. Current research mainly focuses on model merging…

Crowded in B-Space: Calibrating Shared Directions for LoRA Merging

2026-04-18 · Yixuan Tang, Yi Yang arxiv

Merging separately trained LoRA adapters is a practical alternative to joint multi-task training, but it often hurts performance. Existing methods usually treat the LoRA update $ΔW = BA$ as a single object and do not dis…

LoRA on the Go: Instance-level Dynamic LoRA Selection and Merging

2025-11-10 · Seungeon Lee, Soumi Das, Manish Gupta, Krishna P. Gummadi arxiv

Low-Rank Adaptation (LoRA) has emerged as a parameter-efficient approach for fine-tuning large language models. However, conventional LoRA adapters are typically trained for a single task, limiting their applicability in…

BYOM: Building Your Own Multi-Task Model For Free

2023-10-03 · Weisen Jiang, Baijiong Lin, Han Shi, Yu Zhang 외

Recently, various merging methods have been proposed to build a multi-task model from task-specific finetuned models without retraining. However, existing methods suffer from a large performance deterioration compared to…

Saliency-Aware Model Merging

2026-05-30 · Jungin Park, Jiyoung Lee, Kwanghoon Sohn arxiv

Model merging aims to consolidate multiple task-specific models fine-tuned on different datasets into a unified architecture that performs cross-domain proficiency. Current data-free model merging methods often struggle …

Test-time Adaptation