paper-with-me

홈 › Papers

Syntactic Surprisal From Neural Models Predicts, But Underestimates, Human Processing Difficulty From Syntactic Ambiguities

2022-10-21 · Suhas Arehalli, Brian Dillon, Tal Linzen

Humans exhibit garden path effects: When reading sentences that are temporarily structurally ambiguous, they slow down when the structure is disambiguated in favor of the less preferred alternative. Surprisal theory (Hale, 2001; Levy, 2008), a prominent explanation of this finding, proposes that these slowdowns are due to the unpredictability of each of the words that occur in these sentences. Challenging this hypothesis, van Schijndel & Linzen (2021) find that estimates of the cost of word predictability derived from language models severely underestimate the magnitude of human garden path effects. In this work, we consider whether this underestimation is due to the fact that humans weight syntactic factors in their predictions more highly than language models do. We propose a method for estimating syntactic predictability from a language model, allowing us to weigh the cost of lexical and syntactic predictability independently. We find that treating syntactic predictability independently from lexical predictability indeed results in larger estimates of garden path. At the same time, even when syntactic predictability is independently weighted, surprisal still greatly underestimate the magnitude of human garden path effects. Our results support the hypothesis that predictability is not the only factor responsible for the processing cost associated with garden path sentences.

📄 PDF Abstract BibTeX arXiv:2210.12187

Code (1)

sarehalli/syntacticsurprisal 공식 구현 pytorch

Tasks

Language Modelling

Similar Papers 제목 키워드 기반

An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal

2026-04-20 · Ryo Yoshida, Shinnosuke Isono, Taiga Someya, Yohei Oseki 외 arxiv

Surprisal theory hypothesizes that the difficulty of human sentence processing increases linearly with surprisal, the negative log-probability of a word given its context. Computational psycholinguistics has tested this …

Syntactic Belief Update as the Driver of Garden Path Processing Difficulty

2026-06-25 · Alan Zhou, Miloš Stanojević, John T. Hale arxiv

Garden path sentences present a processing difficulty for humans -- the sentence prefix leads the listener towards one interpretation, until the listener hears a critical word that shows that the initial interpretation w…

Dual Alignment Between Language Model Layers and Human Sentence Processing

2026-04-20 · Tatsuki Kuribayashi, Alex Warstadt, Yohei Oseki, Ethan Gotlieb Wilcox arxiv

A recent study (Kuribayashi et al., 2025) has shown that human sentence processing behavior, typically measured on syntactically unchallenging constructions, can be effectively modeled using surprisal from early layers o…

Why are language models less surprised than humans? Testing the Parse Multiplicity Mismatch Hypothesis

2026-05-14 · William Timkey, Brian Dillon, Tal Linzen arxiv

Surprisal theory posits that the processing difficulty of a word is determined by its predictability in context, offering a potential link between human sentence processing and next-word predictions from language models.…

False perspectives on human language: why statistics needs linguistics

2023-02-17 · Matteo Greco, Andrea Cometa, Fiorenzo Artoni, Robert Frank 외

A sharp tension exists about the nature of human language between two opposite parties: those who believe that statistical surface distributions, in particular using measures like surprisal, provide a better understandin…