paper-with-me

홈 › Papers

MAGIC: Achieving Superior Model Merging via Magnitude Calibration

2025-12-22 · Yayuan Li, Jian Zhang, Jintao Guo, Zihan Cheng, Lei Qi, Yinghuan Shi, Yang Gao arxiv

The proliferation of pre-trained models has given rise to a wide array of specialised, fine-tuned models. Model merging aims to merge the distinct capabilities of these specialised models into a unified model, requiring minimal or even no additional training. A core objective of model merging is to ensure the merged model retains the behavioural characteristics of the specialised models, typically achieved through feature alignment. We identify that features consist of two critical components: direction and magnitude. Prior research has predominantly focused on directional alignment, while the influence of magnitude remains largely neglected, despite its pronounced vulnerability to perturbations introduced by common merging operations (e.g., parameter fusion and sparsification). Such perturbations to magnitude inevitably lead to feature deviations in the merged model from the specialised models, resulting in subsequent performance degradation. To address this, we propose MAGnItude Calibration (MAGIC), a plug-and-play framework that rectifies layer-wise magnitudes in feature and weight spaces, with three variants. Specifically, our Feature Space Calibration (FSC) realigns the merged model's features using a small set of unlabelled data, while Weight Space Calibration (WSC) extends this calibration to the weight space without requiring additional data. Combining these yields Dual Space Calibration (DSC). Comprehensive experiments demonstrate that MAGIC consistently boosts performance across diverse Computer Vision tasks (+4.3% on eight datasets) and NLP tasks (+8.0% on Llama) without additional training. Our code is available at: https://github.com/lyymuwu/MAGIC

📄 PDF Abstract BibTeX arXiv:2512.19320

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

One-Shot Manipulation Strategy Learning by Making Contact Analogies

2024-11-14 · Yuyao Liu, Jiayuan Mao, Joshua Tenenbaum, Tomás Lozano-Pérez 외

We present a novel approach, MAGIC (manipulation analogies for generalizable intelligent contacts), for one-shot learning of manipulation strategies with fast and extensive generalization to novel objects. By leveraging …

One-Shot Learning

3D Display Calibration by Visual Pattern Analysis

2016-06-23 · Hyoseok Hwang, Hyun Sung Chang, Dongkyung Nam, In So Kweon

Nearly all 3D displays need calibration for correct rendering. More often than not, the optical elements in a 3D display are misaligned from the designed parameter setting. As a result, 3D magic does not perform well as …

MagicAgent: Towards Generalized Agent Planning

2026-02-22 · Xuhui Ren, Shaokang Dong, Chen Yang, Qing Gao 외 arxiv

The evolution of Large Language Models (LLMs) from passive text processors to autonomous agents has established planning as a core component of modern intelligence. However, achieving generalized planning remains elusive…

Reinforcement Learning

A Multimodal Adaptive Graph-based Intelligent Classification Model for Fake News

2024-11-09 · Jun-hao, Xu

Numerous studies have been proposed to detect fake news focusing on multi-modalities based on machine and/or deep learning. However, studies focusing on graph-based structures using geometric deep learning are lacking. T…

Deep LearningFake News DetectionGraph Attention

A Machine Learning Imaging Core using Separable FIR-IIR Filters

2020-01-02 · Masayoshi Asama, Leo F. Isikdogan, Sushma Rao, Bhavin V. Nayak 외

We propose fixed-function neural network hardware that is designed to perform pixel-to-pixel image transformations in a highly efficient way. We use a fully trainable, fixed-topology neural network to build a model that …

BIG-bench Machine LearningColorizationDeblurringDenoising+1