paper-with-me

Papers

Surprisal Theory is Tautological (without Rational Grounding)

2026-07-23 · Ryan Cotterell arxiv

Surprisal theory holds that the human processing difficulty of a linguistic unit in context is an affine function of its surprisal under some language model. I argue this claim is a tautology without further constraint: for any non-negative difficulty measure over units in context, there exists a language model whose surprisal is an affine function of it under mild technical conditions. Therefore, because any pattern of difficulty is consistent with some language model, without an additional constraint on the language model, surprisal theory makes no falsifiable predictions. The tautology was long obscured by an assumption implicit in two decades of psycholinguistic work---that the relevant language model is the distribution that generated the training corpus, so that improving corpus fit improves predictions of human behavior. Recent empirical work has undermined this assumption, demonstrating that better corpus models can be worse predictors of processing difficulty. I conclude that breaking the tautology requires a rationalist intervention, i.e., the relevant language model must be derived from a non-empirically motivated model of the comprehender, which could be based on, for instance, memory constraints or processing goals, and that, thus, does not depend on the behavioral data surprisal theory is meant to explain.

📄 PDF Abstract BibTeX arXiv:2607.21574

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards a Similarity-adjusted Surprisal Theory

2024-10-23 · Clara Meister, Mario Giulianelli, Tiago Pimentel

Surprisal theory posits that the cognitive effort required to comprehend a word is determined by its contextual predictability, quantified as surprisal. Traditionally, surprisal theory treats words as distinct entities, …

Diversity

Testing the Predictions of Surprisal Theory in 11 Languages

2023-07-07 · Ethan Gotlieb Wilcox, Tiago Pimentel, Clara Meister, Ryan Cotterell 외

A fundamental result in psycholinguistics is that less predictable words take a longer time to process. One theoretical explanation for this finding is Surprisal Theory (Hale, 2001; Levy, 2008), which quantifies a word's…

The Generation-Recognition Asymmetry: Six Dimensions of a Fundamental Divide in Formal Language Theory

2026-03-10 · Romain Peyrichou arxiv

Every formal grammar defines a language and can in principle be used in three ways: to generate strings (production), to recognize them (parsing), or -- given only examples -- to infer the grammar itself (grammar inducti…

Dependency Locality and Neural Surprisal as Predictors of Processing Difficulty: Evidence from Reading Times

2021-06-01 · NAACL (CMCL) 2021 6 · Neil Rathi

This paper compares two influential theories of processing difficulty: Gibson (2000)’s Dependency Locality Theory (DLT) and Hale (2001)’s Surprisal Theory. While prior work has aimed to compare DLT and Surprisal Theory (…

surprisal is Not a Theory

2026-07-22 · Andrés Buxó-Lugo, Aniello De Santo, Morgan Grobol, Ryan J. Hubbard 외 arxiv

Surprisal Theory is often characterized as a computational-level explanation per (Marr, 1982). We argue in this work that, even though a computational level narrative has been used to support "representation-agnostic res…