paper-with-me

Papers

Provable Tempered Overfitting of Minimal Nets and Typical Nets

2024-10-24 · Itamar Harel, William M. Hoza, Gal Vardi, Itay Evron, Nathan Srebro, Daniel Soudry

We study the overfitting behavior of fully connected deep Neural Networks (NNs) with binary weights fitted to perfectly classify a noisy training set. We consider interpolation using both the smallest NN (having the minimal number of weights) and a random interpolating NN. For both learning rules, we prove overfitting is tempered. Our analysis rests on a new bound on the size of a threshold circuit consistent with a partial function. To the best of our knowledge, ours are the first theoretical results on benign or tempered overfitting that: (1) apply to deep NNs, and (2) do not require a very high or very low input dimension.

📄 PDF Abstract BibTeX arXiv:2410.19092

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Benign, Tempered, or Catastrophic: A Taxonomy of Overfitting

2022-07-14 · Neil Mallinar, James B. Simon, Amirhesam Abedsoltan, Parthe Pandit 외

The practical success of overparameterized neural networks has motivated the recent scientific study of interpolating methods, which perfectly fit their training data. Certain interpolating methods, including neural netw…

Learning Theory

Interpolation Learning With Minimum Description Length

2023-02-14 · Naren Sarayu Manoj, Nathan Srebro

We prove that the Minimum Description Length learning rule exhibits tempered overfitting. We obtain tempered agnostic finite sample learning guarantees and characterize the asymptotic behavior in the presence of random l…

From Tempered to Benign Overfitting in ReLU Neural Networks

2023-05-24 · NeurIPS 2023 11 · Guy Kornowski, Gilad Yehudai, Ohad Shamir

Overparameterized neural networks (NNs) are observed to generalize well even when trained to perfectly fit noisy data. This phenomenon motivated a large body of work on "benign overfitting", where interpolating predictor…

Characterizing Overfitting in Kernel Ridgeless Regression Through the Eigenspectrum

2024-02-02 · Tin Sum Cheng, Aurelien Lucchi, Anastasis Kratsios, David Belius

We derive new bounds for the condition number of kernel matrices, which we then use to enhance existing non-asymptotic test error bounds for kernel ridgeless regression (KRR) in the over-parameterized regime for a fixed …

regression

Noisy Interpolation Learning with Shallow Univariate ReLU Networks

2023-07-28 · Nirmit Joshi, Gal Vardi, Nathan Srebro

Understanding how overparameterized neural networks generalize despite perfect interpolation of noisy training data is a fundamental question. Mallinar et. al. 2022 noted that neural networks seem to often exhibit ``temp…

regression