paper-with-me

홈 › Papers

Where Should LoRA Go? Component-Type Placement in Hybrid Language Models

2026-04-24 · Hector Borobia, Elies Seguí-Mas, Guillermina Tormo-Carbó arxiv

Hybrid language models that interleave attention with recurrent components are increasingly competitive with pure Transformers, yet standard LoRA practice applies adapters uniformly without considering the distinct functional roles of each component type. We systematically study component-type LoRA placement across two hybrid architectures -- Qwen3.5-0.8B (sequential, GatedDeltaNet + softmax attention) and Falcon-H1-0.5B (parallel, Mamba-2 SSM + attention) -- fine-tuned on three domains and evaluated on five benchmarks. We find that the attention pathway -- despite being the minority component -- consistently outperforms full-model adaptation with 5-10x fewer trainable parameters. Crucially, adapting the recurrent backbone is destructive in sequential hybrids (-14.8 pp on GSM8K) but constructive in parallel ones (+8.6 pp). We further document a transfer asymmetry: parallel hybrids exhibit positive cross-task transfer while sequential hybrids suffer catastrophic forgetting. These results establish that hybrid topology fundamentally determines adaptation response, and that component-aware LoRA placement is a necessary design dimension for hybrid architectures.

📄 PDF Abstract BibTeX arXiv:2604.22127

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

PLoP: Precise LoRA Placement for Efficient Finetuning of Large Models

2025-06-25 · Soufiane Hayou, Nikhil Ghosh, Bin Yu

Low-Rank Adaptation (LoRA) is a widely used finetuning method for large models. Its small memory footprint allows practitioners to adapt large models to specific tasks at a fraction of the cost of full finetuning. Differ…

One Step Beyond: Feedthrough & Placement-Aware Rectilinear Floorplanner

2025-07-20 · Zhexuan Xu, Jie Wang, Siyuan Xu, Zijie Geng 외 arxiv

Floorplanning determines the shapes and locations of modules on a chip canvas and plays a critical role in optimizing the chip's Power, Performance, and Area (PPA) metrics. However, existing floorplanning approaches ofte…

Component Centric Placement Using Deep Reinforcement Learning

2026-02-26 · Kart Leong Lim arxiv

Automated placement of components on printed circuit boards (PCBs) is a critical stage in placement layout design. While reinforcement learning (RL) has been successfully applied to system-on-chip IP block placement and …

Reinforcement Learning

Replacement Policy of Systems with Dependent Components via Integration of Dynamic Programming and Simulated Annealing

2019-07-24

In a dependent multi-component system, increasing the deterioration of a part leads to the increased deterioration rate of other parts as well. In these systems, a deterioration limit is usually pre-determined for each p…

Inadequacy of Linear Methods for Minimal Sensor Placement and Feature Selection in Nonlinear Systems; a New Approach Using Secants

2021-01-27 · Samuel E. Otto, Clarence W. Rowley

Sensor placement and feature selection are critical steps in engineering, modeling, and data science that share a common mathematical theme: the selected measurements should enable solution of an inverse problem. Most re…

feature selection