paper-with-me

홈 › Papers

Tracing Computation Density in LLMs

2026-05-26 · Corentin Kervadec, Iuliia Lysova, Iuri Macocco, Marco Baroni, Gemma Boleda arxiv

Transformer-based large language models (LLMs) are comprised of billions of parameters arranged in deep and wide computational graphs, but it is not clear that they exploit their full capacity for all inputs. We introduce the s-Trace method to efficiently estimate a subgraph of size s that approximates a full model output. With this method, we find the computation in a variety of LLMs to be organized in two distinct phases. A small subgraph mostly composed of early-layer nodes can reconstruct the head of the full model output distribution. Adding further nodes, mostly located in later layers and increasingly consisting of attention heads, leads to incremental refinements in approximating the full output distribution. We find moreover that the amount of necessary computation per input correlates with model uncertainty, and that sparser subgraphs encode shallow statistics, such as unigram frequency. Overall, our results suggest a consistent modular organization in effective LLM computation, with a sparse early-layer core providing a rough prediction that is further refined through denser computations in later layers.

📄 PDF Abstract BibTeX arXiv:2605.27033

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MemDefrag: Latent Memory Defragmentation for Large Language Models

2026-07-07 · Ruiyi Yan, Zhuoyuan Mao, Yiwen Guo arxiv

Latent memory, which stores past knowledge fragments as per-layer hidden states, has emerged as a promising paradigm (e.g., MemoryLLM and M+) for long-term memory in large language models (LLMs). However, the paradigm su…

Sparse or Dense? A Mechanistic Estimation of Computation Density in Transformer-based LLMs

2026-01-30 · Corentin Kervadec, Iuliia Lysova, Marco Baroni, Gemma Boleda arxiv

Transformer-based large language models (LLMs) are comprised of billions of parameters arranged in deep and wide computational graphs. Several studies on LLM efficiency optimization argue that it is possible to prune a s…

PureLight: Learning Complex Luminaires with Light Tracing

2026-06-03 · Pedro Figueiredo, Zixuan Li, Beibei Wang, Miloš Hašan 외 arxiv

We propose a neural formulation for estimating the appearance of complex luminaires. We focus on challenging luminaires with complex light transport (e.g., small emitters enclosed by multiple specular layers) that are di…

Exploring the Potential of Large Language Models in Generating Code-Tracing Questions for Introductory Programming Courses

2023-10-23 · Aysa Xuemo Fan, Ranran Haoran Zhang, Luc Paquette, Rui Zhang

In this paper, we explore the application of large language models (LLMs) for generating code-tracing questions in introductory programming courses. We designed targeted prompts for GPT4, guiding it to generate code-trac…

Tracing and segmentation of molecular patterns in 3-dimensional cryo-et/em density maps through algorithmic image processing and deep learning-based techniques

2024-03-26 · Salim Sazzed

Understanding the structures of biological macromolecules is highly important as they are closely associated with cellular functionalities. Comprehending the precise organization actin filaments is crucial because they f…

Electron Tomography