paper-with-me

홈 › Papers

How Well Does First-Token Entropy Approximate Word Entropy as a Psycholinguistic Predictor?

2025-07-29 · Christian Clark, Byung-Doh Oh, William Schuler arxiv

Contextual entropy is a psycholinguistic measure capturing the anticipated difficulty of processing a word just before it is encountered. Recent studies have tested for entropy-related effects as a potential complement to well-known effects from surprisal. For convenience, entropy is typically estimated based on a language model's probability distribution over a word's first subword token. However, this approximation results in underestimation and potential distortion of true word entropy. To address this, we generate Monte Carlo (MC) estimates of word entropy that allow words to span a variable number of tokens. Regression experiments on reading times show divergent results between first-token and MC word entropy, suggesting a need for caution in using first-token approximations of contextual entropy.

📄 PDF Abstract BibTeX arXiv:2507.22209

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

When Does Multi-Agent Collaboration Help? An Entropy Perspective

2026-02-04 · Yuxuan Zhao, Sijia Chen, Ningxin Su arxiv

Multi-agent systems (MAS) have emerged as a prominent paradigm for leveraging large language models (LLMs) to tackle complex tasks. However, the mechanisms governing the effectiveness of MAS built upon publicly available…

Distribution Prompting: Understanding the Expressivity of Language Models Through the Next-Token Distributions They Can Produce

2025-05-18 · Haojin Wang, Zining Zhu, Freda Shi

Autoregressive neural language models (LMs) generate a probability distribution over tokens at each time step given a prompt. In this work, we attempt to systematically understand the probability distributions that LMs c…

GMTS: Gradient Magnitude-based Token Selection Improves RLVR Training for LLM Reasoning

2026-08-31 · Outongyi Lv, Yuanwei Zhang, Xiaoqun Zhang arxiv

Reinforcement learning (RL), particularly RL with Verifiable Rewards (RLVR), has recently emerged as a central paradigm for enhancing large language models' (LLMs) reasoning abilities, demonstrating remarkable effectiven…

Reinforcement Learning

DiffAdapt: Difficulty-Adaptive Reasoning for Token-Efficient LLM Inference

2025-10-22 · Xiang Liu, Xuming Hu, Xiaowen Chu, Eunsol Choi arxiv

Recent reasoning Large Language Models (LLMs) demonstrate remarkable problem-solving abilities but often generate long thinking traces whose utility is unclear. Our work aims to improve their efficiency, enabling them to…

Sequential KV Cache Compression via Probabilistic Language Tries: Beyond the Per-Vector Shannon Limit

2026-04-10 · Gregory Magarshak arxiv

Recent work on KV cache quantization, culminating in TurboQuant, has approached the Shannon entropy limit for per-vector compression of transformer key-value caches. We observe that this limit applies to a strictly weake…