paper-with-me

홈 › Papers

Diversity in Large Language Models under Supervised Fine-Tuning

2026-04-30 · Roman Klypa, Oleksandr Cherednichenko arxiv

Supervised Fine-Tuning (SFT) is essential for aligning Large Language Models (LLMs) with user intent, yet it is believed to suppress generative diversity. Although this reduction is frequently referenced, formal empirical testing of the phenomenon remains limited. The expressiveness of LLMs by itself was addressed by multiple prior methods. Their varying perspectives suggest that deeper investigation could yield further improvements. In this study, we attribute the decline to two primary drivers: the neglect of low-frequency patterns within fine-tuning datasets and the forgetting of preexisting knowledge. Motivated by our theoretical analysis, we develop Tempered Focal (TOFU) loss, a novel objective that addresses both stated challenges simultaneously. Our extensive evaluation confirms at scale that generation breadth narrows after SFT and strengthens the hypothesis explaining this effect. Across multiple models and benchmarks, we demonstrate that TOFU enhances output diversity while preserving high response quality, offering a principled approach to SFT.

📄 PDF Abstract BibTeX arXiv:2605.00195

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Effect of Language Diversity When Fine-Tuning Large Language Models for Translation

2025-05-19 · David Stap, Christof Monz

Prior research diverges on language diversity in LLM fine-tuning: Some studies report benefits while others find no advantages. Through controlled fine-tuning experiments across 132 translation directions, we systematica…

DiversityTranslation

On the Diversity of Synthetic Data and its Impact on Training Large Language Models

2024-10-19 · Hao Chen, Abdul Waheed, Xiang Li, Yidong Wang 외

The rise of Large Language Models (LLMs) has accentuated the need for diverse, high-quality pre-training data. Synthetic data emerges as a viable solution to the challenges of data scarcity and inaccessibility. While pre…

Diversity

Measuring diversity of synthetic prompts and data generated with fine-grained persona prompting

2025-05-23 · Gauri Kambhatla, Chantal Shaib, Venkata Govindarajan

Fine-grained personas have recently been used for generating 'diverse' synthetic data for pre-training and supervised fine-tuning of Large Language Models (LLMs). In this work, we measure the diversity of persona-driven …

Diversity

#InsTag: Instruction Tagging for Analyzing Supervised Fine-tuning of Large Language Models

2023-08-14 · Keming Lu, Hongyi Yuan, Zheng Yuan, Runji Lin 외

Foundation language models obtain the instruction-following ability through supervised fine-tuning (SFT). Diversity and complexity are considered critical factors of a successful SFT dataset, while their definitions rema…

DiversityInstruction FollowingTAG

From Macro to Micro: Probing Dataset Diversity in Language Model Fine-Tuning

2025-05-30 · Haoyu Li, Xuhong LI, Yiming Dong, Kun Liu

Dataset diversity plays a pivotal role for the successful training of many machine learning models, particularly in the supervised fine-tuning (SFT) stage of large language model (LLM) development. Despite increasing rec…

DiversityLanguage ModelingLanguage ModellingLarge Language Model