paper-with-me

Papers

Modeling intra-textual variation with entropy and surprisal: topical vs. stylistic patterns

2017-08-01 · WS 2017 8 · Stefania Degaetano-Ortlieb, Elke Teich

We present a data-driven approach to investigate intra-textual variation by combining entropy and surprisal. With this approach we detect linguistic variation based on phrasal lexico-grammatical patterns across sections of research articles. Entropy is used to detect patterns typical of specific sections. Surprisal is used to differentiate between more and less informationally-loaded patterns as well as type of information (topical vs. stylistic). While we here focus on research articles in biology/genetics, the methodology is especially interesting for digital humanities scholars, as it can be applied to any text type or domain and combined with additional variables (e.g. time, author or social group).

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Articles

Similar Papers 제목 키워드 기반

On the Effect of Anticipation on Reading Times

2022-11-25 · Tiago Pimentel, Clara Meister, Ethan G. Wilcox, Roger Levy 외

Over the past two decades, numerous studies have demonstrated how less predictable (i.e., higher surprisal) words take more time to read. In general, these studies have implicitly assumed the reading process is purely re…

STARE: Surprisal-Guided Token-Level Advantage Reweighting for Policy Entropy Stability

2026-06-17 · Haipeng Luo, Qingfeng Sun, Songli Wu, Can Xu 외 arxiv

Reinforcement Learning with Verifiable Rewards algorithms like GRPO have emerged as the dominant post-training paradigm for complex reasoning in LLMs, yet commonly suffer from policy entropy collapse during training. We …

Reinforcement Learning

A Modeling Study of the Effects of Surprisal and Entropy in Perceptual Decision Making of an Adaptive Agent

2019-06-01 · WS 2019 6 · Pyeong Whan Cho, Richard Lewis

Processing difficulty in online language comprehension has been explained in terms of surprisal and entropy reduction. Although both hypotheses have been supported by experimental data, we do not fully understand their r…

Decision Making

Contextual Semantic Relevance and Word Surprisal Predict N400 and P600 Dynamics During Naturalistic Reading

2026-07-05 · Kun Sun, Rong Wang arxiv

Word surprisal is a well-established computational predictor of human neural responses during language comprehension, but it remains less clear whether local semantic fit explains neural response variation beyond lexical…

Testing the Predictions of Surprisal Theory in 11 Languages

2023-07-07 · Ethan Gotlieb Wilcox, Tiago Pimentel, Clara Meister, Ryan Cotterell 외

A fundamental result in psycholinguistics is that less predictable words take a longer time to process. One theoretical explanation for this finding is Surprisal Theory (Hale, 2001; Levy, 2008), which quantifies a word's…