paper-with-me

홈 › Papers

An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal

2026-04-20 · Ryo Yoshida, Shinnosuke Isono, Taiga Someya, Yohei Oseki, Tatsuki Kuribayashi arxiv

Surprisal theory hypothesizes that the difficulty of human sentence processing increases linearly with surprisal, the negative log-probability of a word given its context. Computational psycholinguistics has tested this hypothesis using language models (LMs) as proxies for human prediction. While surprisal derived from recent neural LMs generally captures human processing difficulty on naturalistic corpora that predominantly consist of simple sentences, it severely underestimates processing difficulty on sentences that require syntactic disambiguation (garden-path effects). This leads to the claim that the processing difficulty of such sentences cannot be reduced to surprisal, although it remains possible that neural LMs simply differ from humans in next-word prediction. In this paper, we investigate whether it is truly impossible to construct a neural LM that can explain garden-path effects via surprisal. Specifically, instead of evaluating off-the-shelf neural LMs, we fine-tune these LMs on garden-path sentences so as to better align surprisal-based reading-time estimates with actual human reading times. Our results show that fine-tuned LMs do not overfit and successfully capture human reading slowdowns on held-out garden-path items; they even improve predictive power for human reading times on naturalistic corpora and preserve their general LM capabilities. These results provide an existence proof for a neural LM that can explain both garden-path effects and naturalistic reading times via surprisal, but also raise a theoretical question: what kind of evidence can truly falsify surprisal theory?

📄 PDF Abstract BibTeX arXiv:2604.18293

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

When the LM misunderstood the human chuckled: Analyzing garden path effects in humans and language models

2025-02-13 · Samuel Joseph Amouyal, Aya Meltzer-Asscher, Jonathan Berant

Modern Large Language Models (LLMs) have shown human-like abilities in many language tasks, sparking interest in comparing LLMs' and humans' language processing. In this paper, we conduct a detailed comparison of the two…

Image GenerationSentenceText to Image GenerationText-to-Image Generation

ERAS: Evaluating the Robustness of Chinese NLP Models to Morphological Garden Path Errors

2024-10-16 · Qinchan Li, Sophie Hao

In languages without orthographic word boundaries, NLP models perform word segmentation, either as an explicit preprocessing step or as an implicit step in an end-to-end computation. This paper shows that Chinese NLP mod…

SegmentationSentenceSentiment Analysis

Syntactic Surprisal From Neural Models Predicts, But Underestimates, Human Processing Difficulty From Syntactic Ambiguities

2022-10-21 · Suhas Arehalli, Brian Dillon, Tal Linzen

Humans exhibit garden path effects: When reading sentences that are temporarily structurally ambiguous, they slow down when the structure is disambiguated in favor of the less preferred alternative. Surprisal theory (Hal…

Language Modelling

Incremental Comprehension of Garden-Path Sentences by Large Language Models: Semantic Interpretation, Syntactic Re-Analysis, and Attention

2024-05-25 · Andrew Li, Xianle Feng, Siddhant Narang, Austin Peng 외

When reading temporarily ambiguous garden-path sentences, misinterpretations sometimes linger past the point of disambiguation. This phenomenon has traditionally been studied in psycholinguistic experiments using online …

Question AnsweringSentenceTask 2

The role of inhibitory control in garden-path sentence processing: A Chinese-English bilingual perspective

2024-12-13 · Xiaohui Rao, Haoze Li, Xiaofang Lin, Lijuan Liang

In reading garden-path sentences, people must resolve competing interpretations, though initial misinterpretations can linger despite reanalysis. This study examines the role of inhibitory control (IC) in managing these …

Sentence