paper-with-me

홈 › Papers

Uniform convergence may be unable to explain generalization in deep learning

2019-02-13 · NeurIPS 2019 12 · Vaishnavh Nagarajan, J. Zico Kolter

Aimed at explaining the surprisingly good generalization behavior of overparameterized deep networks, recent works have developed a variety of generalization bounds for deep learning, all based on the fundamental learning-theoretic technique of uniform convergence. While it is well-known that many of these existing bounds are numerically large, through numerous experiments, we bring to light a more concerning aspect of these bounds: in practice, these bounds can {\em increase} with the training dataset size. Guided by our observations, we then present examples of overparameterized linear classifiers and neural networks trained by gradient descent (GD) where uniform convergence provably cannot "explain generalization" -- even if we take into account the implicit bias of GD {\em to the fullest extent possible}. More precisely, even if we consider only the set of classifiers output by GD, which have test errors less than some small $\epsilon$ in our settings, we show that applying (two-sided) uniform convergence on this set of classifiers will yield only a vacuous generalization guarantee larger than $1-\epsilon$. Through these findings, we cast doubt on the power of uniform convergence-based generalization bounds to provide a complete picture of why overparameterized deep networks generalize well.

📄 PDF Abstract BibTeX arXiv:1902.04742

Code (1)

locuslab/uniform-convergence-NeurIPS19 공식 구현 tf

Tasks

Deep LearningGeneralization Bounds

Similar Papers 제목 키워드 기반

Generalization Error Bounds of Gradient Descent for Learning Over-parameterized Deep ReLU Networks

2019-02-04 · Yuan Cao, Quanquan Gu

Empirical studies show that gradient-based methods can learn deep neural networks (DNNs) with very good generalization performance in the over-parameterization regime, where DNNs can easily fit a random labeling of the t…

Generalization Bounds

Towards Understanding Generalization via Decomposing Excess Risk Dynamics

2021-06-11 · ICLR 2022 4 · Jiaye Teng, Jianhao Ma, Yang Yuan

Generalization is one of the fundamental issues in machine learning. However, traditional techniques like uniform convergence may be unable to explain generalization under overparameterization. As alternative approaches,…

Generalization Bounds

Explaining generalization in deep learning: progress and fundamental limits

2021-10-17 · Vaishnavh Nagarajan

This dissertation studies a fundamental open challenge in deep learning theory: why do deep networks generalize well even while being overparameterized, unregularized and fitting the training data to zero error? In the f…

Deep LearningGeneralization BoundsLearning Theory

On Uniform Convergence and Low-Norm Interpolation Learning

2020-06-10 · NeurIPS 2020 12 · Lijia Zhou, Danica J. Sutherland, Nathan Srebro

We consider an underdetermined noisy linear regression model where the minimum-norm interpolating predictor is known to be consistent, and ask: can uniform convergence in a norm ball, or at least (following Nagarajan and…

Uniform Convergence of Interpolators: Gaussian Width, Norm Bounds, and Benign Overfitting

2021-06-17 · NeurIPS 2021 12 · Frederic Koehler, Lijia Zhou, Danica J. Sutherland, Nathan Srebro

We consider interpolation learning in high-dimensional linear regression with Gaussian data, and prove a generic uniform convergence guarantee on the generalization error of interpolators in an arbitrary hypothesis class…

Generalization Boundsregression