paper-with-me

홈 › Papers

Eigen-Spike Emergence and Quadratic Equivalents for Conjugate Kernels on Nonlinearly Separable Data

2026-05-28 · Collin Cranston, Zhichao Wang, Todd Kemp, Michael W. Mahoney arxiv

Recent work in random matrix theory (RMT) has developed the notion of deterministic equivalents: typically linear surrogate models that approximate the spectral behavior of large nonlinear random matrices, such as nonlinear feature maps in neural networks (NNs). Such equivalents make theoretical predictions tractable by reducing a complex model to a simpler one with properties that fall under the umbrella of classical RMT tools. However, this leaves open the question of whether this idealized linear equivalence remains meaningful for classification of high-dimensional nonlinearly separable data. Motivated by this, we consider the conjugate kernel (CK), which is the nonlinear feature map of a one-layer feedforward NN, under a canonical nonlinearly separable dataset for the XOR problem; and we use the study of informative outlier eigenvalues in the CK and whether their corresponding eigenvectors asymptotically align with XOR labels as a proxy for nonlinear learnability. We develop a robust quadratic equivalent of the CK matrix that enables a precise analysis of emergent informative spikes, as one modifies various knobs common in ML practice: sample complexity, signal-to-noise ratio (SNR), nonlinear activation choice, and pretrained features. We identify regimes in which these knobs move the CK beyond the linear equivalent and produce BBP-type transitions to label-aligned outlier eigenspaces. Our analysis helps bring deterministic-equivalence tools from RMT to bear on problems of practical relevance in ML.

📄 PDF Abstract BibTeX arXiv:2605.29669

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Nonlinear spiked covariance matrices and signal propagation in deep neural networks

2024-02-15 · Zhichao Wang, Denny Wu, Zhou Fan

Many recent works have studied the eigenvalue spectrum of the Conjugate Kernel (CK) defined by the nonlinear feature map of a feedforward neural network. However, existing results only establish weak convergence of the e…

Representation Learning

Asymptotic Theory of Eigenvectors for Latent Embeddings with Generalized Laplacian Matrices

2025-03-01 · Jianqing Fan, Yingying Fan, Jinchi Lv, Fan Yang 외

Laplacian matrices are commonly employed in many real applications, encoding the underlying latent structural information such as graphs and manifolds. The use of the normalization terms naturally gives rise to random ma…

Uncertainty Quantification

Spectral Evolution and Invariance in Linear-width Neural Networks

2022-11-11 · NeurIPS 2023 11

We investigate the spectral properties of linear-width feed-forward neural networks, where the sample size is asymptotically proportional to network width. Empirically, we show that the spectra of weight in this high dim…

Efficient non-conjugate Gaussian process factor models for spike count data using polynomial approximations

2019-06-07 · Stephen L. Keeley, David M. Zoltowski, Yiyi Yu, Jacob L. Yates 외

Gaussian Process Factor Analysis (GPFA) has been broadly applied to the problem of identifying smooth, low-dimensional temporal structure underlying large-scale neural recordings. However, spike trains are non-Gaussian, …

Variational Inference

Efficient non-conjugate Gaussian process factor models for spike countdata using polynomial approximations

2020-01-01 · ICML 2020 1 · Stephen Keeley, David Zoltowski, Jonathan Pillow, Spencer Smith 외

Gaussian Process Factor Analysis (GPFA) hasbeen broadly applied to the problem of identi-fying smooth, low-dimensional temporal struc-ture underlying large-scale neural recordings.However, spike trains are non-Gaussian, …

Variational Inference