paper-with-me

홈 › Papers

Eigen Neural Network: Unlocking Generalizable Vision with Eigenbasis

2025-08-02 · Anzhe Cheng, Chenzhong Yin, Mingxi Cheng, Shukai Duan, Shahin Nazarian, Paul Bogdan arxiv

The remarkable success of Deep Neural Networks(DNN) is driven by gradient-based optimization, yet this process is often undermined by its tendency to produce disordered weight structures, which harms feature clarity and degrades learning dynamics. To address this fundamental representational flaw, we introduced the Eigen Neural Network (ENN), a novel architecture that reparameterizes each layer's weights in a layer-shared, learned orthonormal eigenbasis. This design enforces decorrelated, well-aligned weight dynamics axiomatically, rather than through regularization, leading to more structured and discriminative feature representations. When integrated with standard BP, ENN consistently outperforms state-of-the-art methods on large-scale image classification benchmarks, including ImageNet, and its superior representations generalize to set a new benchmark in cross-modal image-text retrieval. Furthermore, ENN's principled structure enables a highly efficient, backpropagation-free(BP-free) local learning variant, ENN-$\ell$. This variant not only resolves BP's procedural bottlenecks to achieve over 2$\times$ training speedup via parallelism, but also, remarkably, surpasses the accuracy of end-to-end backpropagation. ENN thus presents a new architectural paradigm that directly remedies the representational deficiencies of BP, leading to enhanced performance and enabling a more efficient, parallelizable training regime.

📄 PDF Abstract BibTeX arXiv:2508.01219

Code (0)

등록된 구현이 없습니다.

Tasks

Image ClassificationText Retrieval

Similar Papers 제목 키워드 기반

Graph Distillation with Eigenbasis Matching

2023-10-13 · Yang Liu, Deyu Bo, Chuan Shi

The increasing amount of graph data places requirements on the efficient training of graph neural networks (GNNs). The emerging graph distillation (GD) tackles this challenge by distilling a small synthetic graph to repl…

Blind Deconvolution of Graph Signals: Robustness to Graph Perturbations

2024-12-19 · Chang Ye, Gonzalo Mateos

We study blind deconvolution of signals defined on the nodes of an undirected graph. Although observations are bilinear functions of both unknowns, namely the forward convolutional filter coefficients and the graph signa…

Denoising

Fast Linear Reservoirs via Diagonalization

2026-02-23 · Romain de Coudenhove, Yannis Bendi-Ouis, Anthony Strock, Xavier Hinaut arxiv

We introduce a diagonalization-based optimization for Linear Echo State Networks (ESNs) that reduces the per-step computational complexity of reservoir state updates from quadratic to linear. By reformulating reservoir d…

EMoE: Eigenbasis-Guided Routing for Mixture-of-Experts

2026-01-17 · Anzhe Cheng, Shukai Duan, Shixuan Li, Chenzhong Yin 외 arxiv

The relentless scaling of deep learning models has led to unsustainable computational demands, positioning Mixture-of-Experts (MoE) architectures as a promising path towards greater efficiency. However, MoE models are pl…

Eigenvalue-corrected Natural Gradient Based on a New Approximation

2020-11-27 · Kai-Xin Gao, Xiao-Lei Liu, Zheng-Hai Huang, Min Wang 외

Using second-order optimization methods for training deep neural networks (DNNs) has attracted many researchers. A recently proposed method, Eigenvalue-corrected Kronecker Factorization (EKFAC) (George et al., 2018), pro…