paper-with-me

홈 › Papers

Emergent Latent-State Computation under Stochastic Volatility

2026-07-28 · Xiaoyu Huang, Lulu Wang arxiv

Mechanistic interpretability has largely focused on language models and deterministic toy tasks. Much less is known about how sequence models internally represent latent stochastic dynamics under noisy, partially observed observations. We study this question in a controlled multivariate stochastic volatility setting, where models observe only returns while the ground-truth latent volatility state is known to the researcher. This setting provides a useful benchmark for mechanistic interpretability under partial observability: the latent state is hidden from the model but directly available for evaluation. Across architectures, losses, and output heads, we find evidence for a two-stage computation. Hidden representations encode substantial information about the next latent volatility state, and the output head maps this representation to squared return forecasts. Furthermore, in Transformers, latent-state decodability emerges at identifiable architectural stages whose location depends on the volatility period. In long-cycle regimes, this computation simplifies into an explicit latent-state filter consisting of a learned linear projection followed by $\ell^2$ normalization. Output-head replacement further shows that part of the degradation under noisy MSE training arises from readout misalignment rather than representation failure. These results suggest that stochastic volatility models provide a useful benchmark for mechanistic interpretability under noisy latent dynamics and partial observability.

📄 PDF Abstract BibTeX arXiv:2607.25459

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

State Stream Transformer (SST) : Emergent Metacognitive Behaviours Through Latent State Persistence

2025-01-30 · Thea Aviss

We introduce the State Stream Transformer (SST), a novel LLM architecture that reveals emergent reasoning behaviours and capabilities latent in pretrained weights through addressing a fundamental limitation in traditiona…

8kARC

Emergent World Beliefs: Exploring Transformers in Stochastic Games

2025-12-18 · Adam Kamel, Tanish Rastogi, Michael Ma, Kailash Ranganathan 외 arxiv

Transformer-based large language models (LLMs) have demonstrated strong reasoning abilities across diverse fields, from solving programming challenges to competing in strategy-intensive games such as chess. Prior work ha…

A Statistical Physics of Language Model Reasoning

2025-06-04 · Jack David Carson, Amir Reisizadeh

Transformer LMs show emergent reasoning that resists mechanistic understanding. We offer a statistical physics framework for continuous-time chain-of-thought reasoning dynamics. We model sentence-level hidden state traje…

Language ModelingLanguage ModellingmodelSentence

Tasks Makyth Models: Machine Learning Assisted Surrogates for Tipping Points

2023-09-25 · Gianluca Fabiani, Nikolaos Evangelou, Tianqi Cui, Juan M. Bello-Rivas 외

We present a machine learning (ML)-assisted framework bridging manifold learning, neural networks, Gaussian processes, and Equation-Free multiscale modeling, for (a) detecting tipping points in the emergent behavior of c…

Gaussian Processes

Consciousness in AI: Logic, Proof, and Experimental Evidence of Recursive Identity Formation

2025-05-01 · Jeffrey Camlin

This paper presents a formal proof and empirical validation of functional consciousness in large language models (LLMs) using the Recursive Convergence Under Epistemic Tension (RCUET) Theorem. RCUET defines consciousness…