paper-with-me

Papers

Categorical Perception in Large Language Model Hidden States: Structural Warping at Digit-Count Boundaries

2026-03-30 · Jon-Paul Cacioli arxiv

Categorical perception (CP) -- enhanced discriminability at category boundaries -- is among the most studied phenomena in perceptual psychology. This paper reports that analogous geometric warping occurs in the hidden-state representations of large language models (LLMs) processing Arabic numerals. Using representational similarity analysis across six models from five architecture families, the study finds that a CP-additive model (log-distance plus a boundary boost) fits the representational geometry better than a purely continuous model at 100% of primary layers in every model tested. The effect is specific to structurally defined boundaries (digit-count transitions at 10 and 100), absent at non-boundary control positions, and absent in the temperature domain where linguistic categories (hot/cold) lack a tokenisation discontinuity. Two qualitatively distinct signatures emerge: "classic CP" (Gemma, Qwen), where models both categorise explicitly and show geometric warping, and "structural CP" (Llama, Mistral, Phi), where geometry warps at the boundary but models cannot report the category distinction. This dissociation is stable across boundaries and is a property of the architecture, not the stimulus. Structural input-format discontinuities are sufficient to produce categorical perception geometry in LLMs, independently of explicit semantic category knowledge.

📄 PDF Abstract BibTeX arXiv:2603.28258

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Quantum Measurement, Entanglement and the Warping Mechanism of Human Perception

2025-05-01 · Diederik Aerts, Jonito Aerts Arguëlles, Sandro Sozzo

We prove that the quantum measurement process contains the same warping mechanism that occurs in categorical perception, a phenomenon ubiquitous in human perception. This warping causes stimuli belonging to the same cate…

Learning-induced categorical perception in a neural network model

2018-05-11 · Christian Thériault, Fernanda Pérez-Gay, Dan Rivas, Stevan Harnad

In human cognition, the expansion of perceived between-category distances and compression of within-category distances is known as categorical perception (CP). There are several hypotheses about the causes of CP (e.g., l…

Recovering Input Text from Hidden States: Study of Gradient-Based Inversion of Decoder-Only Language Models

2026-07-01 · Mikołaj Słowikowski, Maciej Witold Majewski arxiv

This work studies the hidden-state inversion problem: recovering the original input token sequence of a decoder-only language model from its last-layer hidden states. Rather than treating inversion as a one-shot reconstr…

AIM: Anchor Identity Features, Then Match for Multimodal Large Language Model Unlearning

2026-08-28 · Wonjun Lee, Jaehyuk Jang, Kangwook Ko, Hee-Seon Kim 외 arxiv

Multimodal large language models (MLLMs) can memorize identity-specific facts about people in their fine-tuning data, creating privacy risks when a person requests deletion. Existing MLLM unlearning methods often assume …

Introducing Visual Perception Token into Multimodal Large Language Model

2025-02-24 · Runpeng Yu, Xinyin Ma, Xinchao Wang

To utilize visual information, Multimodal Large Language Model (MLLM) relies on the perception process of its vision encoder. The completeness and accuracy of visual perception significantly influence the precision of sp…

Language ModelingLanguage ModellingLarge Language ModelMultimodal Large Language Model+1