paper-with-me

홈 › Papers

Escaping the Mode Lottery: Multi-Response Training Improves Language Model Generalization

2026-05-30 · Hasan Amin, Kian Ahrabian, Ming Yin, Rajiv Khanna arxiv

Modern language-model fine-tuning typically pairs each prompt with a single response, even though many prompts admit multiple valid completions. This effectively reduces a multi-modal conditional distribution to a one-sample view, a phenomenon we call the "mode lottery," where training emphasizes a subset of plausible modes while leaving others underrepresented. We study multi-response training (MRT), which retains multiple responses per prompt, and develop a principled account of when and why it helps. Our key insight is that prompts and responses are distinct statistical resources: additional prompts reduce uncertainty about the input distribution, while additional responses reduce uncertainty about the conditional output distribution. This yields a variance-budget tradeoff that predicts when retaining multiple responses is worthwhile, shows diminishing returns as prompt-level uncertainty dominates, and explains why large redundant corpora can exhibit an implicit multi-response effect. We further analyze response selection, and show that Random-K-of-N is the unbiased default for distributional fine-tuning, reward-based selection can induce mode collapse, and a submodular quality-diversity objective provides an efficient alternative with theoretical guarantees. Controlled simulations validate the predicted variance and selection effects, including a striking failure mode where reward-only selection produces gradients misaligned with the true objective. Across structured and real-world datasets, including a new multi-prompt, multi-response benchmark, MRT consistently improves distributional generalization, with the largest gains in high response-diversity, low prompt-redundancy regimes. MRT reframes response multiplicity as a data-allocation problem with clear guidance: when responses are cheap and diverse, keeping more than one is not a heuristic, but a statistically grounded choice.

📄 PDF Abstract BibTeX arXiv:2606.00544

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Gender gaps in frontier entrepreneurship? Evidence from 1901 Oklahoma land lottery winners

2022-06-29 · Jason Poulos

The paper investigates gender differences in entrepreneurship by exploiting a large-scale land lottery in Oklahoma at the turn of the 20$^{\text{th}}$ century. Lottery winners claimed land in the order in which their nam…

Liquidity Constraints, Cash Windfalls, and Entrepreneurship: Evidence from Administrative Data on Lottery Winners

2023-03-29 · Hsuan-Hua Huang, Hsing-Wen Han, Kuang-Ta Lo, Tzu-Ting Yang

Using administrative data on Taiwanese lottery winners, this paper examines the effects of cash windfalls on entrepreneurship. We compare the start-up decisions of households winning more than 1.5 million NTD (50,000 USD…

LotteryFL: Personalized and Communication-Efficient Federated Learning with Lottery Ticket Hypothesis on Non-IID Datasets

2020-08-07 · Ang Li, Jingwei Sun, Binghui Wang, Lin Duan 외

Federated learning is a popular distributed machine learning paradigm with enhanced privacy. Its primary goal is learning a global model that offers good performance for the participants as many as possible. The technolo…

Federated Learning

Lottery Jackpots Exist in Pre-trained Models

2021-04-18 · Yuxin Zhang, Mingbao Lin, Yunshan Zhong, Fei Chao 외

Network pruning is an effective approach to reduce network complexity with acceptable performance compromise. Existing studies achieve the sparsity of neural networks via time-consuming weight training or complex searchi…

Network Pruning

KS-Lottery: Finding Certified Lottery Tickets for Multilingual Language Models

2024-02-05 · Fei Yuan, Chang Ma, Shuai Yuan, Qiushi Sun 외

The lottery ticket hypothesis posits the existence of ``winning tickets'' within a randomly initialized neural network. Do winning tickets exist for LLMs in fine-tuning scenarios? How can we find such winning tickets? In…

Translation