paper-with-me

홈 › Papers

When Context Misleads: Surprisal, Energy and Attention Entropy as Metrics of Coherence Illusions in LLMs

2026-06-19 · Ece Takmaz, Nitin Kumar, Li Kloostra, Jakub Dotlacil arxiv

Psycholinguistics studies show that human readers fall for coherence illusions: an incoherent discourse can seem coherent simply because a distractor matches what comes next. We investigate whether Dutch language models (6 monolingual and 4 multilingual) show the same behavior on texts that link back to earlier context with words such as 'again' and 'too'. First, we find that surprisal at the critical word tracks human acceptability judgments and eye-tracking data. Models are more surprised by incoherent continuations, but a matching distractor in the prior context reduces this surprisal. Second, attention entropy at the critical position identifies heads that behave differently under coherence vs. incoherence. We find that ablating these heads shows transfer effects across experiments, suggesting a shared mechanism. Third, we introduce energy from the associative-memory literature as a metric to quantify discourse coherence. Taken together, our results show that coherence illusions arise in Dutch LLMs, with entropy and energy exposing mechanisms that operate across settings.

📄 PDF Abstract BibTeX arXiv:2606.21203

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Energy-Based Transformers as Predictors of Reading Difficulty

2026-06-22 · Jakub Dotlacil, Ece Takmaz arxiv

Transformer language models have become established tools for modeling human sentence processing, with measures such as surprisal and attention entropy serving as effective predictors of reading difficulty that together …

On the Role of Context in Reading Time Prediction

2024-09-12 · Andreas Opedal, Eleanor Chodroff, Ryan Cotterell, Ethan Gotlieb Wilcox

We present a new perspective on how readers integrate context during real-time language comprehension. Our proposals build on surprisal theory, which posits that the processing effort of a linguistic unit (e.g., a word) …

Language ModelingLanguage ModellingPrediction

Towards a Similarity-adjusted Surprisal Theory

2024-10-23 · Clara Meister, Mario Giulianelli, Tiago Pimentel

Surprisal theory posits that the cognitive effort required to comprehend a word is determined by its contextual predictability, quantified as surprisal. Traditionally, surprisal theory treats words as distinct entities, …

Diversity

The Effect of Surprisal on Reading Times in Information Seeking and Repeated Reading

2024-10-10 · Keren Gruteke Klein, Yoav Meiri, Omer Shubi, Yevgeni Berzak

The effect of surprisal on processing difficulty has been a central topic of investigation in psycholinguistics. Here, we use eyetracking data to examine three language processing regimes that are common in daily life bu…

The grip of grammar on meaning uncertainty: cross-linguistic evidence, neural correlates, and clinical relevance

2026-05-02 · Rui He, Claudio Palominos, Samuele Vallisa, Ni Yang 외 arxiv

Isolated word meanings are inherently uncertain. This uncertainty reduces when they are combined and anchored in context. We propose that grammar compresses meaning uncertainty cross-linguistically, which is reflected in…