paper-with-me

홈 › Papers

Reinforced Model Merging

2025-03-27 · Jiaqi Han, Jingwen Ye, Shunyu Liu, Haofei Zhang, Jie Song, Zunlei Feng, Mingli Song

The success of large language models has garnered widespread attention for model merging techniques, especially training-free methods which combine model capabilities within the parameter space. However, two challenges remain: (1) uniform treatment of all parameters leads to performance degradation; (2) search-based algorithms are often inefficient. In this paper, we present an innovative framework termed Reinforced Model Merging (RMM), which encompasses an environment and agent tailored for merging tasks. These components interact to execute layer-wise merging actions, aiming to search the optimal merging architecture. Notably, RMM operates without any gradient computations on the original models, rendering it feasible for edge devices. Furthermore, by utilizing data subsets during the evaluation process, we addressed the bottleneck in the reward feedback phase, thereby accelerating RMM by up to 100 times. Extensive experiments demonstrate that RMM achieves state-of-the-art performance across various vision and NLP datasets and effectively overcomes the limitations of the existing baseline methods. Our code is available at https://github.com/WuDiHJQ/Reinforced-Model-Merging.

📄 PDF Abstract BibTeX arXiv:2503.21272

Code (1)

wudihjq/reinforced-model-merging 공식 구현 pytorch

Tasks

model

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

Behavior Knowledge Merge in Reinforced Agentic Models

2026-01-20 · Xiangchi Yuan, Dachuan Shi, Chunhui Zhang, Zheyuan Liu 외 arxiv

Reinforcement learning (RL) is central to post-training, particularly for agentic models that require specialized reasoning behaviors. In this setting, model merging offers a practical mechanism for integrating multiple …

Reinforcement Learning

Reinforced Continual Learning

2018-05-31 · NeurIPS 2018 12 · Ju Xu, Zhanxing Zhu

Most artificial intelligence models have limiting ability to solve new tasks faster, without forgetting previously acquired knowledge. The recently emerging paradigm of continual learning aims to solve this issue, in whi…

Continual LearningGeneral Classificationreinforcement-learningReinforcement Learning+1

Learning to Selectively Transfer: Reinforced Transfer Learning for Deep Text Matching

2018-12-30 · Chen Qu, Feng Ji, Minghui Qiu, Liu Yang 외

Deep text matching approaches have been widely studied for many applications including question answering and information retrieval systems. To deal with a domain that has insufficient labeled data, these approaches can …

Information RetrievalNatural Language InferenceParaphrase IdentificationQuestion Answering+3

Instance Segmentation of Fibers from Low Resolution CT Scans via 3D Deep Embedding Learning

2019-01-04 · Tomasz Konopczyński, Thorben Kröger, Lei Zheng, Jürgen Hesser

We propose a novel approach for automatic extraction (instance segmentation) of fibers from low resolution 3D X-ray computed tomography scans of short glass fiber reinforced polymers. We have designed a 3D instance segme…

3D Instance SegmentationClusteringInstance SegmentationSegmentation+1

AutoFS: Automated Feature Selection via Diversity-aware Interactive Reinforcement Learning

2020-08-27 · Wei Fan, Kunpeng Liu, Hao liu, Pengyang Wang 외

In this paper, we study the problem of balancing effectiveness and efficiency in automated feature selection. Feature selection is a fundamental intelligence for machine learning and predictive analysis. After exploring …

Diversityfeature selectionNavigatereinforcement-learning+2