A Continuum of Generation Tasks for Investigating Length Bias and Degenerate Repetition
Language models suffer from various degenerate behaviors. These differ between tasks: machine translation (MT) exhibits length bias, while tasks like story generation exhibit excessive repetition. Recent work has attributed the difference to task constrainedness, but evidence for this claim has always involved many confounding variables. To study this question directly, we introduce a new experimental framework that allows us to smoothly vary task constrainedness, from MT at one end to fully open-ended generation at the other, while keeping all other aspects fixed. We find that: (1) repetition decreases smoothly with constrainedness, explaining the difference in repetition across tasks; (2) length bias surprisingly also decreases with constrainedness, suggesting some other cause for the difference in length bias; (3) across the board, these problems affect the mode, not the whole distribution; (4) the differences cannot be attributed to a change in the entropy of the distribution, since another method of changing the entropy, label smoothing, does not produce the same effect.
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationStory GenerationSimilar Papers 제목 키워드 기반
MoSS: Monocular Shape Sensing for Continuum Robots
Continuum robots are promising candidates for interactive tasks in medical and industrial applications due to their unique shape, compliance, and miniaturization capability. Accurate and real-time shape sensing is essent…
DecoderOn Uncertainties in Cross-Correlation Lags and the Reality of Wavelength-Dependent Continuum Lags in Active Galactic Nuclei
We describe a model-independent method of assessing the uncertainties in cross-correlation lags determined from AGN light curves, and use this method to investigate the reality of lags between UV and optical continuum va…
A Long Way to Go: Investigating Length Correlations in RLHF
Great success has been reported using Reinforcement Learning from Human Feedback (RLHF) to align large language models, with open preference datasets enabling wider experimentation, particularly for "helpfulness" in task…
Question AnsweringWhat they do when in doubt: a study of inductive biases in seq2seq learners
Sequence-to-sequence (seq2seq) learners are widely used, but we still have only limited knowledge about what inductive biases shape the way they generalize. We address that by investigating how popular seq2seq learners g…
MemorizationInvestigating Societal Biases in a Poetry Composition System
There is a growing collection of work analyzing and mitigating societal biases in language understanding, generation, and retrieval tasks, though examining biases in creative tasks remains underexplored. Creative languag…
Data AugmentationRetrievalStyle Transfer