paper-with-me

홈 › Papers

Reinforcement Learning Fine-Tunes a Sparse Subnetwork in Large Language Models

2025-07-23 · Andrii Balashov arxiv

Reinforcement learning (RL) is a key post-pretraining step for aligning large language models (LLMs) with complex tasks and human preferences. While it is often assumed that RL fine-tuning requires updating most of a model's parameters, we challenge this assumption with a surprising finding: RL fine-tuning consistently modifies only a small subnetwork (typically 5-30% of weights), leaving most parameters unchanged. We call this phenomenon RL-induced parameter update sparsity. It arises naturally, without any sparsity constraints or parameter-efficient tuning, and appears across multiple RL algorithms (e.g., PPO, DPO, SimPO, PRIME) and model families (e.g., OpenAI, Meta, and open-source LLMs). Moreover, the subnetworks updated by RL show substantial overlap across different seeds, datasets, and algorithms-far exceeding chance-suggesting a partially transferable structure in the pretrained model. We show that fine-tuning only this sparse subnetwork recovers full model performance and yields parameters nearly identical to the fully fine-tuned model. Our analysis suggests this sparsity emerges because RL operates near the model's original distribution, requiring only targeted changes. KL penalties, gradient clipping, and on-policy dynamics have limited effect on the sparsity pattern. These findings shed new light on how RL adapts models: not by shifting all weights, but by focusing training on a small, consistently updated subnetwork. This insight enables more efficient RL methods and reframes sparsity through the lens of the lottery ticket hypothesis.

📄 PDF Abstract BibTeX arXiv:2507.17107

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Reinforcement Learning Finetunes Small Subnetworks in Large Language Models

2025-05-16 · Sagnik Mukherjee, Lifan Yuan, Dilek Hakkani-Tur, Hao Peng

Reinforcement learning (RL) yields substantial improvements in large language models (LLMs) downstream task performance and alignment with human values. Surprisingly, such large gains result from updating only a small su…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

PARP: Prune, Adjust and Re-Prune for Self-Supervised Speech Recognition

2021-06-10 · NeurIPS 2021 12 · Cheng-I Jeff Lai, Yang Zhang, Alexander H. Liu, Shiyu Chang 외

Self-supervised speech representation learning (speech SSL) has demonstrated the benefit of scale in learning rich representations for Automatic Speech Recognition (ASR) with limited paired data, such as wav2vec 2.0. We …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Representation LearningSelf-Supervised Learning+3

Sparse Subnetwork Enhancement for Underrepresented Languages in Large Language Models

2025-10-15 · Daniil Gurgurov, Tanja Baeumel, Josef van Genabith, Simon Ostermann arxiv

Large language models (LLMs) exhibit substantial performance disparities across languages, particularly between high- and low-resource settings. We propose a framework for improving performance in underrepresented langua…

The Multiple Ticket Hypothesis: Random Sparse Subnetworks Suffice for RLVR

2026-02-02 · Israel Adewuyi, Solomon Okibe, Vladmir Ivanov arxiv

The Lottery Ticket Hypothesis demonstrated that sparse subnetworks can match full-model performance, suggesting parameter redundancy. Meanwhile, in Reinforcement Learning with Verifiable Rewards (RLVR), recent work has s…

Reinforcement Learning

Diverse Lottery Tickets Boost Ensemble from a Single Pretrained Model

2022-05-24 · BigScience (ACL) 2022 5 · Sosuke Kobayashi, Shun Kiyono, Jun Suzuki, Kentaro Inui

Ensembling is a popular method used to improve performance as a last resort. However, ensembling multiple models finetuned from a single pretrained model has been not very effective; this could be due to the lack of dive…

Diversity