paper-with-me

홈 › Papers

Global Low-Rank, Local Full-Rank: The Holographic Encoding of Learned Algorithms

2026-02-20 · Yongzhong Xu arxiv

Grokking -- the abrupt transition from memorization to generalization after extended training -- has been linked to the emergence of low-dimensional structure in learning dynamics. Yet neural network parameters inhabit extremely high-dimensional spaces. How can a low-dimensional learning process produce solutions that resist low-dimensional compression? We investigate this question in multi-task modular arithmetic, training shared-trunk Transformers with separate heads for addition, multiplication, and a quadratic operation modulo 97. Across three model scales (315K--2.2M parameters) and five weight decay settings, we compare three reconstruction methods: per-matrix SVD, joint cross-matrix SVD, and trajectory PCA. Across all conditions, grokking trajectories are confined to a 2--6 dimensional global subspace, while individual weight matrices remain effectively full-rank. Reconstruction from 3--5 trajectory PCs recovers over 95\% of final accuracy, whereas both per-matrix and joint SVD fail at sub-full rank. Even when static decompositions capture most spectral energy, they destroy task-relevant structure. These results show that learned algorithms are encoded through dynamically coordinated updates spanning all matrices, rather than localized low-rank components. We term this the holographic encoding principle: grokked solutions are globally low-rank in the space of learning directions but locally full-rank in parameter space, with implications for compression, interpretability, and understanding how neural networks encode computation.

📄 PDF Abstract BibTeX arXiv:2602.18649

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Learning to Rank Question Answer Pairs with Holographic Dual LSTM Architecture

2017-07-20 · Tay Yi, Phan Minh C., Tuan Luu Anh, Hui Siu Cheung

We describe a new deep learning architecture for learning to rank question answer pairs. Our approach extends the long short-term memory (LSTM) network with holographic composition to model the relationship between quest…

Feature EngineeringLearning-To-RankSentence

Fully Parallel Architecture for Semi-global Stereo Matching with Refined Rank Method

2019-05-07 · Yiwu Yao, Yuhua Cheng

Fully parallel architecture at disparity-level for efficient semi-global matching (SGM) with refined rank method is presented. The improved SGM algorithm is implemented with the non-parametric unified rank model which is…

Stereo MatchingStereo Matching Hand

Global-to-Local or Local-to-Global? Enhancing Image Retrieval with Efficient Local Search and Effective Global Re-ranking

2025-09-04 · Dror Aiger, Bingyi Cao, Kaifeng Chen, Andre Araujo arxiv

The dominant paradigm in image retrieval systems today is to search large databases using global image features, and re-rank those initial results with local image feature matching techniques. This design, dubbed global-…

Computational EfficiencyImage Retrieval

Non-local Meets Global: An Iterative Paradigm for Hyperspectral Image Restoration

2020-10-24 · wei he, Quanming Yao, Chao Li, Naoto Yokoya 외

Non-local low-rank tensor approximation has been developed as a state-of-the-art method for hyperspectral image (HSI) restoration, which includes the tasks of denoising, compressed HSI reconstruction and inpainting. Unfo…

DenoisingImage Restoration

Non-local Meets Global: An Integrated Paradigm for Hyperspectral Denoising

2018-12-11 · CVPR 2019 6 · Wei He, Quanming Yao, Chao Li, Naoto Yokoya 외

Non-local low-rank tensor approximation has been developed as a state-of-the-art method for hyperspectral image (HSI) denoising. Unfortunately, with more spectral bands for HSI, while the running time of these methods si…

DenoisingHyperspectral Image Denoising