paper-with-me

홈 › Papers

Geometric Routing Enables Causal Expert Control in Mixture of Experts

2026-04-15 · Ivan Ternovtsii, Yurii Bilak arxiv

Sparse Mixture-of-Experts (MoE) models scale parameters while fixing active computation per token, but the specialization of individual experts remains opaque. In a companion paper we showed that routing topology is quality-neutral: five structurally different configurations converge to statistically equivalent language modeling quality. Here we show that expert identity is nonetheless causally meaningful: individual rank-1 experts are monosemantic by construction, and cosine-similarity routing in a low-dimensional metric space makes their specialization directly inspectable. We present four lines of evidence. First, projecting expert output vectors through the unembedding matrix yields a Semantic Dictionary: 15% of experts are monosemantic specialists spanning 10 categories (temporal, geographic, cardinal, discourse, emotional, financial, military, scientific). Second, routing exhibits a frequency-to-syntax gradient: early layers separate tokens by word frequency, deeper layers by syntactic class (Zipf-confound controls, all $p < 0.001$). Third, causal interventions confirm these labels: steering toward a temporal expert's centroid increases P(temporal) by +321% (median across 44 prompts); suppressing a geographic expert drops P(geographic) by -23%; rewriting an expert's output vector halves target-category probability, and effects compose additively across layers. Fourth, the interventions are not unique to cosine routing: linear routers support comparable steering, but only cosine routing provides geometric transparency -- expert specialization is readable directly from the centroid matrix. MoE expert-level specialization is a first-class interpretability primitive: architecturally monosemantic, causally validated, and controllable at inference with zero overhead.

📄 PDF Abstract BibTeX arXiv:2604.14434

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Grassmannian Mixture-of-Experts: Concentration-Controlled Routing on Subspace Manifolds

2026-02-19 · Ibne Farabi Shihab, Sanjeda Akter, Anuj Sharma arxiv

Mixture-of-Experts models rely on learned routers to assign tokens to experts, yet standard softmax gating provides no principled mechanism to control the tradeoff between sparsity and utilization. We propose Grassmannia…

Equifinality in Mixture of Experts: Routing Topology Does Not Determine Language Modeling Quality

2026-04-15 · Ivan Ternovtsii, Yurii Bilak arxiv

Sparse Mixture-of-Experts (MoE) architectures employ increasingly sophisticated routing mechanisms -- learned routers, multi-hop trajectories, token-dependent gating. We ask: does routing topology actually determine lang…

Geometric Asymmetry in MoE Specialization: Functional Decorrelation and Representational Overlap

2026-05-08 · Feilong Liu arxiv

Mixture-of-Experts (MoE) architectures achieve scalable capacity through sparse routing, yet the geometric structure of expert specialization remains poorly understood. We introduce a unified Jacobian-PCA-Grassmann frame…

Beyond Geometric Complementarity: Coherent Overlap in Sparse Mixture-of-Experts Routing

2026-07-30 · Huiyuan Tian, Bonan Xu, Shijian Li arxiv

Sparse mixture-of-experts (MoE) language models route each token to multiple experts, suggesting a geometric account of their benefit: co-selected experts should contribute distinct representation directions. Existing ev…

Polysemantic Experts, Monosemantic Paths: Routing as Control in MoEs

2026-04-20 · Charles Ye, Bo Yuan, Lee Sharkey arxiv

An LLM's residual stream is both state and instruction: it encodes the current context and determines the next transformation. We introduce a parameter-free decomposition for Mixture-of-Experts models that splits each la…