paper-with-me

홈 › Papers

Towards Initialization-dependent and Non-vacuous Generalization Bounds for Overparameterized Shallow Neural Networks

2026-04-01 · Yunwen Lei, Yufeng Xie arxiv

Overparameterized neural networks often show a benign overfitting property in the sense of achieving excellent generalization behavior despite the number of parameters exceeding the number of training examples. A promising direction to explain benign overfitting is to relate generalization to the norm of distance from initialization, motivated by the empirical observations that this distance is often significantly smaller than the norm itself. However, the existing initialization-dependent complexity analyses measure the distance from initialization by the Frobenius norm, and often imply vacuous bounds in practice for overparamterized models. In this paper, we develop initialization-dependent complexity bounds for shallow neural networks with general Lipschitz activation functions. Our bounds depend on the path-norm of the distance from initialization, which are derived by introducing a new peeling technique to handle the challenge along with the initialization-dependent constraint. We also develop a lower bound tight up to a constant factor. Finally, we conduct empirical comparisons and show that our generalization analysis implies non-vacuous bounds for overparameterized networks.

📄 PDF Abstract BibTeX arXiv:2604.00505

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Survey on Data-Dependent Worst-Case Generalization Bounds

2026-05-13 · Hubert Leroux, Jean Marcus, Julien Roger arxiv

Deep neural networks generalize well despite being heavily overparameterized, in apparent contradiction with classical learning theory based on uniform convergence over fixed hypothesis spaces. Uniform bounds over the en…

Uniform Generalization Bounds for Overparameterized Neural Networks

2021-09-13 · Sattar Vakili, Michael Bromberg, Jezabel Garcia, Da-Shan Shiu 외

An interesting observation in artificial neural networks is their favorable generalization error despite typically being extremely overparameterized. It is well known that the classical statistical learning methods often…

Generalization Bounds

Uniform convergence may be unable to explain generalization in deep learning

2019-02-13 · NeurIPS 2019 12 · Vaishnavh Nagarajan, J. Zico Kolter

Aimed at explaining the surprisingly good generalization behavior of overparameterized deep networks, recent works have developed a variety of generalization bounds for deep learning, all based on the fundamental learnin…

Deep LearningGeneralization Bounds

Explaining generalization in deep learning: progress and fundamental limits

2021-10-17 · Vaishnavh Nagarajan

This dissertation studies a fundamental open challenge in deep learning theory: why do deep networks generalize well even while being overparameterized, unregularized and fitting the training data to zero error? In the f…

Deep LearningGeneralization BoundsLearning Theory

Non-Vacuous Generalization Bounds at the ImageNet Scale: A PAC-Bayesian Compression Approach

2018-04-16 · ICLR 2019 5 · Wenda Zhou, Victor Veitch, Morgane Austern, Ryan P. Adams 외

Modern neural networks are highly overparameterized, with capacity to substantially overfit to training data. Nevertheless, these networks often generalize well in practice. It has also been observed that trained network…

Generalization Bounds