paper-with-me

Papers

Zero Generalization Error Theorem for Random Interpolators via Algebraic Geometry

2025-12-06 · Naoki Yoshida, Isao Ishikawa, Masaaki Imaizumi arxiv

We theoretically demonstrate that the generalization error of interpolators for machine learning models under teacher-student settings becomes 0 once the number of training samples exceeds a certain threshold. Understanding the high generalization ability of large-scale models such as deep neural networks (DNNs) remains one of the central open problems in machine learning theory. While recent theoretical studies have attributed this phenomenon to the implicit bias of stochastic gradient descent (SGD) toward well-generalizing solutions, empirical evidences indicate that it primarily stems from properties of the model itself. Specifically, even randomly sampled interpolators, which are parameters that achieve zero training error, have been observed to generalize effectively. In this study, under a teacher-student framework, we prove that the generalization error of randomly sampled interpolators becomes exactly zero once the number of training samples exceeds a threshold determined by the geometric structure of the interpolator set in parameter space. As a proof technique, we leverage tools from algebraic geometry to mathematically characterize this geometric structure.

📄 PDF Abstract BibTeX arXiv:2512.06347

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Exact Gap between Generalization Error and Uniform Convergence in Random Feature Models

2021-03-08 · Zitong Yang, Yu Bai, Song Mei

Recent work showed that there could be a large gap between the classical uniform convergence bound and the actual test error of zero-training-error predictors (interpolators) such as deep neural networks. To better under…

Uniform Convergence of Interpolators: Gaussian Width, Norm Bounds, and Benign Overfitting

2021-06-17 · NeurIPS 2021 12 · Frederic Koehler, Lijia Zhou, Danica J. Sutherland, Nathan Srebro

We consider interpolation learning in high-dimensional linear regression with Gaussian data, and prove a generic uniform convergence guarantee on the generalization error of interpolators in an arbitrary hypothesis class…

Generalization Boundsregression

Uniform Convergence of Interpolators: Gaussian Width, Norm Bounds and Benign Overfitting

2021-05-21 · NeurIPS 2021 12 · Frederic Koehler, Lijia Zhou, Danica J. Sutherland, Nathan Srebro

We consider interpolation learning in high-dimensional linear regression with Gaussian data, and prove a generic uniform convergence guarantee on the generalization error of interpolators in an arbitrary hypothesis class…

Generalization Boundsregression

Flatness and Generalization: Learning Multi-Index Models with Homogeneous Neural Networks

2026-06-03 · Harsh Vardhan, Hossein Taheri, Arya Mazumdar arxiv

A common heuristic used to explain the generalization of first-order gradient methods on non-convex neural networks is that "flat interpolators generalize well" (Hochreiter and Schmidhuber, 1994; Keskar et al., 2017), wh…

Precise analysis of ridge interpolators under heavy correlations -- a Random Duality Theory view

2024-06-13 · Mihailo Stojnic

We consider fully row/column-correlated linear regression models and study several classical estimators (including minimum norm interpolators (GLS), ordinary least squares (LS), and ridge regressors). We show that \emph{…

FormTime Series