paper-with-me

홈 › Papers

Clozing the Gap: Exploring Why Language Model Surprisal Outperforms Cloze Surprisal

2026-01-14 · Sathvik Nair, Byung-Doh Oh arxiv

How predictable a word is can be quantified in two ways: using human responses to the cloze task or using probabilities from language models (LMs).When used as predictors of processing effort, LM probabilities outperform probabilities derived from cloze data. However, it is important to establish that LM probabilities do so for the right reasons, since different predictors can lead to different scientific conclusions about the role of prediction in language comprehension. We present evidence for three hypotheses about the advantage of LM probabilities: not suffering from low resolution, distinguishing semantically similar words, and accurately assigning probabilities to low-frequency words. These results call for efforts to improve the resolution of cloze studies, coupled with experiments on whether human-like prediction is also as sensitive to the fine-grained distinctions made by LM probabilities.

📄 PDF Abstract BibTeX arXiv:2601.09886

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Modeling the Impact of Syntactic Distance and Surprisal on Cross-Slavic Text Comprehension

2022-06-01 · LREC 2022 6 · Irina Stenger, Philip Georgis, Tania Avgustinova, Bernd Möbius 외

We focus on the syntactic variation and measure syntactic distances between nine Slavic languages (Belarusian, Bulgarian, Croatian, Czech, Polish, Slovak, Slovene, Russian, and Ukrainian) using symmetric measures of inse…

Cloze TestReading Comprehension

Surprisal and Metaphor Novelty Judgments: Moderate Correlations and Divergent Scaling Effects Revealed by Corpus-Based and Synthetic Datasets

2026-01-05 · Omar Momen, Emilie Sitter, Berenike Herrmann, Sina Zarrieß arxiv

Novel metaphor comprehension involves complex semantic processes and linguistic creativity, making it an interesting task for studying language models (LMs). This study investigates whether surprisal, a probabilistic mea…

Generalized Measures of Anticipation and Responsivity in Online Language Processing

2024-09-16 · Mario Giulianelli, Andreas Opedal, Ryan Cotterell

We introduce a generalization of classic information-theoretic measures of predictive uncertainty in online language processing, based on the simulation of expected continuations of incremental linguistic contexts. Our f…

Probabilistic Predictions of People Perusing: Evaluating Metrics of Language Model Performance for Psycholinguistic Modeling

2020-09-08 · EMNLP (CMCL) 2020 11 · Yiding Hao, Simon Mendelsohn, Rachel Sterneck, Randi Martinez 외

By positing a relationship between naturalistic reading times and information-theoretic surprisal, surprisal theory (Hale, 2001; Levy, 2008) provides a natural interface between language models and psycholinguistic model…

Language ModelingLanguage Modelling

CDGP: Automatic Cloze Distractor Generation based on Pre-trained Language Model

2024-03-15 · Shang-Hsuan Chiang, Ssu-Cheng Wang, Yao-Chung Fan

Manually designing cloze test consumes enormous time and efforts. The major challenge lies in wrong option (distractor) selection. Having carefully-design distractors improves the effectiveness of learner ability assessm…

Cloze TestDistractor GenerationLanguage ModelingLanguage Modelling