paper-with-me

홈 › Papers

Flag Varieties: A Geometric Framework for Deep Network Alignment

2026-05-11 · Jingchuan Xiao, Xinyi Sui, Cihan Ruan arxiv

Alignment, the tendency of adjacent weight matrices in deep networks to develop compatible subspace orientations, underlies gradient flow, Neural Collapse, and representation similarity across architectures. Despite extensive empirical documentation, these phenomena have resisted unified theoretical treatment: existing explanations are post-hoc, each fitted to a specific observation with whatever mathematics is at hand. We reverse this direction by deriving the mathematical structure that layerwise alignment inherently demands. Using geometric invariant theory, we prove that alignment geometry has a canonical closed, polystable stratum given by a flag variety, and that subspace intersection dimension is its unique reparameterization-invariant observable, establishing that subspace metrics are not empirical conventions but mathematical necessities. This unified framework yields two dynamical consequences: ridge regularization drives subspace alignment at an exponential rate set by weight decay, whereas nonlinear activations induce a commutator obstruction to exact basis alignment, generically present in nonlinear networks and absent in linear ones. Together these give a geometric explanation of the Level-2/3 hierarchy in Neural Collapse from first principles rather than post-hoc analysis. The commutator magnitude and head subspace overlap further serve as weight-space windows into internal alignment structure, requiring no forward passes. Experiments on multilayer perceptrons, residual networks, and pretrained language models support the proposed diagnostics and delineate their scope.

📄 PDF Abstract BibTeX arXiv:2605.09861

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Semantic Spaces

2016-05-13 · Yuri Manin, Matilde Marcolli

Any natural language can be considered as a tool for producing large databases (consisting of texts, written, or discursive). This tool for its description in turn requires other large databases (dictionaries, grammars e…

When Alignment Hurts: Decoupling Representational Spaces in Multilingual Models

2025-08-18 · Ahmed Elshabrawy, Hour Kaing, Haiyue Song, Alham Fikri Aji 외 arxiv

Alignment with high-resource standard languages is often assumed to aid the modeling of related low-resource varieties. We challenge this assumption by demonstrating that excessive representational entanglement with a do…

AdversariaL attacK sAfety aLIgnment(ALKALI): Safeguarding LLMs through GRACE: Geometric Representation-Aware Contrastive Enhancement- Introducing Adversarial Vulnerability Quality Index (AVQI)

2025-06-10 · Danush Khanna, Krishna Kumar, Basab Ghosh, Vinija Jain 외

Adversarial threats against LLMs are escalating faster than current defenses can adapt. We expose a critical geometric blind spot in alignment: adversarial prompts exploit latent camouflage, embedding perilously close to…

Adversarial AttackSafety Alignment

Phonetic forced alignment for low-resource language varieties: Model training and evaluation on Chengdu Mandarin

2026-07-23 · Zhiheng Qian, Aini Li, Hai Hu, Liang Zhao arxiv

Phonetic forced alignment is a key technique in phonetic research, yet existing alignment systems lack specialized models for low-resource language varieties. We address this by training text-dependent and text-independe…

Skin-Deep: A Geometric Diagnostic for Alignment Fragility in Large Language Model Representations

2026-06-21 · Dongyub Jude Lee, Jungseob Lee, Seungyoon Lee, Seongtae Hong 외 arxiv

Alignment tuning is meant to make harmful-request refusal robust, yet this safety behavior can be erased by a small set of benign fine-tuning examples. This is a deployment risk for open-weight models because a checkpoin…