paper-with-me

Papers

Failure and success of the spectral bias prediction for Kernel Ridge Regression: the case of low-dimensional data

2022-02-07 · Umberto M. Tomasini, Antonio Sclocchi, Matthieu Wyart

Recently, several theories including the replica method made predictions for the generalization error of Kernel Ridge Regression. In some regimes, they predict that the method has a `spectral bias': decomposing the true function $f^*$ on the eigenbasis of the kernel, it fits well the coefficients associated with the O(P) largest eigenvalues, where $P$ is the size of the training set. This prediction works very well on benchmark data sets such as images, yet the assumptions these approaches make on the data are never satisfied in practice. To clarify when the spectral bias prediction holds, we first focus on a one-dimensional model where rigorous results are obtained and then use scaling arguments to generalize and test our findings in higher dimensions. Our predictions include the classification case $f(x)=$sign$(x_1)$ with a data distribution that vanishes at the decision boundary $p(x)\sim x_1^{\chi}$. For $\chi>0$ and a Laplace kernel, we find that (i) there exists a cross-over ridge $\lambda^*_{d,\chi}(P)\sim P^{-\frac{1}{d+\chi}}$ such that for $\lambda\gg \lambda^*_{d,\chi}(P)$, the replica method applies, but not for $\lambda\ll\lambda^*_{d,\chi}(P)$, (ii) in the ridge-less case, spectral bias predicts the correct training curve exponent only in the limit $d\rightarrow\infty$.

📄 PDF Abstract BibTeX arXiv:2202.03348

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Neural tangent kernel analysis of PINN for advection-diffusion equation

2022-11-21 · M. H. Saadat, B. Gjorgiev, L. Das, G. Sansavini

Physics-informed neural networks (PINNs) numerically approximate the solution of a partial differential equation (PDE) by incorporating the residual of the PDE along with its initial/boundary conditions into the loss fun…

Toward a Better Understanding of Fourier Neural Operators from a Spectral Perspective

2024-04-10 · Shaoxiang Qin, Fuyuan Lyu, Wenhui Peng, Dingyang Geng 외

In solving partial differential equations (PDEs), Fourier Neural Operators (FNOs) have exhibited notable effectiveness. However, FNO is observed to be ineffective with large Fourier kernels that parameterize more frequen…

Ensemble Learning

Scaling Continuous Kernels with Sparse Fourier Domain Learning

2024-09-15 · Clayton Harper, Luke Wood, Peter Gerstoft, Eric C. Larson

We address three key challenges in learning continuous kernel representations: computational efficiency, parameter efficiency, and spectral bias. Continuous kernels have shown significant potential, but their practical a…

Computational EfficiencySparse Learning

Scalable Levy Process Priors for Spectral Kernel Learning

2017-12-01 · NeurIPS 2017 12 · Phillip A. Jang, Andrew Loeb, Matthew Davidow, Andrew G. Wilson

Gaussian processes are rich distributions over functions, with generalization properties determined by a kernel function. When used for long-range extrapolation, predictions are particularly sensitive to the choice of ke…

Gaussian Processes

Scalable Lévy Process Priors for Spectral Kernel Learning

2018-02-02 · Phillip A. Jang, Andrew E. Loeb, Matthew B. Davidow, Andrew Gordon Wilson

Gaussian processes are rich distributions over functions, with generalization properties determined by a kernel function. When used for long-range extrapolation, predictions are particularly sensitive to the choice of ke…

Gaussian Processes