paper-with-me

홈 › Papers

Repetitions are not all alike: distinct mechanisms sustain repetition in language models

2025-04-01 · Matéo Mahaut, Francesca Franzon

Text generated by language models (LMs) can degrade into repetitive cycles, where identical word sequences are persistently repeated one after another. Prior research has typically treated repetition as a unitary phenomenon. However, repetitive sequences emerge under diverse tasks and contexts, raising the possibility that it may be driven by multiple underlying factors. Here, we experimentally explore the hypothesis that repetition in LMs can result from distinct mechanisms, reflecting different text generation strategies used by the model. We examine the internal working of LMs under two conditions that prompt repetition: one in which repeated sequences emerge naturally after human-written text, and another where repetition is explicitly induced through an in-context learning (ICL) setup. Our analysis reveals key differences between the two conditions: the model exhibits varying levels of confidence, relies on different attention heads, and shows distinct pattens of change in response to controlled perturbations. These findings suggest that distinct internal mechanisms can interact to drive repetition, with implications for its interpretation and mitigation strategies. More broadly, our results highlight that the same surface behavior in LMs may be sustained by different underlying processes, acting independently or in combination.

📄 PDF Abstract BibTeX arXiv:2504.01100

Code (0)

등록된 구현이 없습니다.

Tasks

AllIn-Context LearningText Generation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

Repetition Facilitates Processing: The Processing Advantage of Construction Repetition in Dialogue

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Repetitions occur frequently in dialogue. This study focuses on the repetition of lexicalised constructions—i.e., recurring multi-word units—in English open domain spoken dialogues. We hypothesise that construction repet…

Language ModelingLanguage Modelling

Analyzing and Mitigating Repetitions in Trip Recommendation

2025-07-26 · Wenzheng Shu, Kangqi Xu, Wenxin Tai, Ting Zhong 외 arxiv

Trip recommendation has emerged as a highly sought-after service over the past decade. Although current studies significantly understand human intention consistency, they struggle with undesired repetitive outcomes that …

Generating Repetitions with Appropriate Repeated Words

2022-07-03 · NAACL 2022 7 · Toshiki Kawamoto, Hidetaka Kamigaito, Kotaro Funakoshi, Manabu Okumura

A repetition is a response that repeats words in the previous speaker's utterance in a dialogue. Repetitions are essential in communication to build trust with others, as investigated in linguistic studies. In this work,…

Language ModelingLanguage Modelling

Learning to Break the Loop: Analyzing and Mitigating Repetitions for Neural Text Generation

2022-06-06 · Jin Xu, Xiaojiang Liu, Jianhao Yan, Deng Cai 외

While large-scale neural language models, such as GPT2 and BART, have achieved impressive results on various text generation tasks, they tend to get stuck in undesirable sentence-level loops with maximization-based decod…

SentenceText GenerationText Summarization

OVR: A Dataset for Open Vocabulary Temporal Repetition Counting in Videos

2024-07-24 · Debidatta Dwibedi, Yusuf Aytar, Jonathan Tompson, Andrew Zisserman

We introduce a dataset of annotations of temporal repetitions in videos. The dataset, OVR (pronounced as over), contains annotations for over 72K videos, with each annotation specifying the number of repetitions, the sta…