paper-with-me

홈 › Papers

A Continuum of Generation Tasks for Investigating Length Bias and Degenerate Repetition

2022-10-19 · Darcey Riley, David Chiang

Language models suffer from various degenerate behaviors. These differ between tasks: machine translation (MT) exhibits length bias, while tasks like story generation exhibit excessive repetition. Recent work has attributed the difference to task constrainedness, but evidence for this claim has always involved many confounding variables. To study this question directly, we introduce a new experimental framework that allows us to smoothly vary task constrainedness, from MT at one end to fully open-ended generation at the other, while keeping all other aspects fixed. We find that: (1) repetition decreases smoothly with constrainedness, explaining the difference in repetition across tasks; (2) length bias surprisingly also decreases with constrainedness, suggesting some other cause for the difference in length bias; (3) across the board, these problems affect the mode, not the whole distribution; (4) the differences cannot be attributed to a change in the entropy of the distribution, since another method of changing the entropy, label smoothing, does not produce the same effect.

📄 PDF Abstract BibTeX arXiv:2210.10817

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationStory Generation

Similar Papers 제목 키워드 기반

MoSS: Monocular Shape Sensing for Continuum Robots

2023-03-02 · Chengnan Shentu, Enxu Li, Chaojun Chen, Puspita Triana Dewi 외

Continuum robots are promising candidates for interactive tasks in medical and industrial applications due to their unique shape, compliance, and miniaturization capability. Accurate and real-time shape sensing is essent…

Decoder

On Uncertainties in Cross-Correlation Lags and the Reality of Wavelength-Dependent Continuum Lags in Active Galactic Nuclei

1998-02-09 · Bradley M. Peterson, Ignaz Wanders, Keith Horne, Stefan Collier 외

We describe a model-independent method of assessing the uncertainties in cross-correlation lags determined from AGN light curves, and use this method to investigate the reality of lags between UV and optical continuum va…

A Long Way to Go: Investigating Length Correlations in RLHF

2023-10-05 · Prasann Singhal, Tanya Goyal, Jiacheng Xu, Greg Durrett

Great success has been reported using Reinforcement Learning from Human Feedback (RLHF) to align large language models, with open preference datasets enabling wider experimentation, particularly for "helpfulness" in task…

Question Answering

What they do when in doubt: a study of inductive biases in seq2seq learners

2020-06-26 · ICLR 2021 1 · Eugene Kharitonov, Rahma Chaabouni

Sequence-to-sequence (seq2seq) learners are widely used, but we still have only limited knowledge about what inductive biases shape the way they generalize. We address that by investigating how popular seq2seq learners g…

Memorization

Investigating Societal Biases in a Poetry Composition System

2020-11-05 · GeBNLP (COLING) 2020 12 · Emily Sheng, David Uthus

There is a growing collection of work analyzing and mitigating societal biases in language understanding, generation, and retrieval tasks, though examining biases in creative tasks remains underexplored. Creative languag…

Data AugmentationRetrievalStyle Transfer