paper-with-me

홈 › Papers

Measuring Dead Directions: Decomposing and Classifying Singular Structure off Canonical Alignment

2026-07-01 · Tejas Pradeep Shirodkar arxiv

We give a descent-free, alignment-free measurement of singular structure on trained networks. At a single frozen checkpoint the read recovers the order $k$ of each dead direction from the directional-Fisher rate, the master invariant from which the per-direction learning coefficient $1/(2k)$ follows exactly, in whatever basis the optimizer left. The same read classifies each direction, separating a genuine singularity, whose order the architecture fixes, from a flat gauge symmetry; the directional-Fisher magnitude settles the cases the order cannot. A pluggable detector supplies the directions for transformer, convolutional, and normalisation layers. The read recovers the architecture-predicted order across constructed cells and trained networks, including a fine-tuned vision transformer whose dead structure is the LayerNorm-kernel gauge and a from-scratch one whose compressed MLP forms a node-death at its activation order. Where the singular structure enumerates, the per-direction orders assemble, through the typed intersection of the loci, into the global coefficient $(λ, m)$ matching the closed form. The method removes the canonical-alignment and descent preconditions of the underlying rate result, turning order-recovery into a deterministic, architecture-general reading. We then map its reach into the Watanabe triple: the order determines the universal singular fluctuation $ν(k)$, though a trained network's realized $ν$ falls below it as the live structure absorbs the dead direction's data fluctuation, and the multiplicity recovers from the dominant structure under a single-locus assumption.

📄 PDF Abstract BibTeX arXiv:2607.00603

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Algebraic Dead Directions in LayerNorm Transformers: A Forward-Pass-Only Diagnostic at LLM Scale

2026-06-17 · Tejas Pradeep Shirodkar, P. J. Narayanan arxiv

Pretrained transformers sit near singular minima of the loss, where the Fisher information metric degenerates along dead directions: directions in parameter space along which the directional Fisher vanishes. Locating suc…

Dead-Direction Signatures: A Cheap Spectral Reading of Singular Complexity

2026-06-19 · Tejas Pradeep Shirodkar, P. J. Narayanan arxiv

Singular learning theory characterises the complexity of a deep network through the geometry of its loss singularities. The local learning coefficient (LLC), the standard estimator of Watanabe's real log canonical thresh…

OPLoRA: Orthogonal Projection LoRA Prevents Catastrophic Forgetting during Parameter-Efficient Fine-Tuning

2025-10-14 · Yifeng Xiong, Xiaohui Xie arxiv

Low-Rank Adaptation (LoRA) enables efficient fine-tuning of large language models but suffers from catastrophic forgetting when learned updates interfere with the dominant singular directions that encode essential pre-tr…

parameter-efficient fine-tuningCode Generation

Dead Directions: Geometric Singular Learning

2026-06-04 · Tejas Pradeep Shirodkar arxiv

Singular learning theory and information geometry study the same spaces: the former in resolved coordinates, the latter in original coordinates under a non-degeneracy assumption that overparameterised models violate. Thi…

Measuring Semantic Similarity by Latent Relational Analysis

2005-08-10 · Peter D. Turney

This paper introduces Latent Relational Analysis (LRA), a method for measuring semantic similarity. LRA measures similarity in the semantic relations between two pairs of words. When two pairs have a high degree of relat…

Multiple-choiceSemantic SimilaritySemantic Textual Similarity