paper-with-me

홈 › Papers

Beyond Benign Overfitting in Nadaraya-Watson Interpolators

2025-02-11 · Daniel Barzilai, Guy Kornowski, Ohad Shamir

In recent years, there has been much interest in understanding the generalization behavior of interpolating predictors, which overfit on noisy training data. Whereas standard analyses are concerned with whether a method is consistent or not, recent observations have shown that even inconsistent predictors can generalize well. In this work, we revisit the classic interpolating Nadaraya-Watson (NW) estimator (also known as Shepard's method), and study its generalization capabilities through this modern viewpoint. In particular, by varying a single bandwidth-like hyperparameter, we prove the existence of multiple overfitting behaviors, ranging non-monotonically from catastrophic, through benign, to tempered. Our results highlight how even classical interpolating methods can exhibit intricate generalization behaviors. In addition, for the purpose of tuning the hyperparameter, the results suggest that over-estimating the intrinsic dimension of the data is less harmful than under-estimating it. Numerical experiments complement our theory, demonstrating the same phenomena.

📄 PDF Abstract BibTeX arXiv:2502.07480

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Uniform Convergence of Interpolators: Gaussian Width, Norm Bounds, and Benign Overfitting

2021-06-17 · NeurIPS 2021 12 · Frederic Koehler, Lijia Zhou, Danica J. Sutherland, Nathan Srebro

We consider interpolation learning in high-dimensional linear regression with Gaussian data, and prove a generic uniform convergence guarantee on the generalization error of interpolators in an arbitrary hypothesis class…

Generalization Boundsregression

Uniform Convergence of Interpolators: Gaussian Width, Norm Bounds and Benign Overfitting

2021-05-21 · NeurIPS 2021 12 · Frederic Koehler, Lijia Zhou, Danica J. Sutherland, Nathan Srebro

We consider interpolation learning in high-dimensional linear regression with Gaussian data, and prove a generic uniform convergence guarantee on the generalization error of interpolators in an arbitrary hypothesis class…

Generalization Boundsregression

Heterogeneous Treatment Effect with Trained Kernels of the Nadaraya-Watson Regression

2022-07-19 · Andrei V. Konstantinov, Stanislav R. Kirpichenko, Lev V. Utkin

A new method for estimating the conditional average treatment effect is proposed in the paper. It is called TNW-CATE (the Trainable Nadaraya-Watson regression for CATE) and based on the assumption that the number of cont…

regressionTransfer Learning

Implicit Regularization Leads to Benign Overfitting for Sparse Linear Regression

2023-02-01 · Mo Zhou, Rong Ge

In deep learning, often the training process finds an interpolator (a solution with 0 training loss), but the test loss is still low. This phenomenon, known as benign overfitting, is a major mystery that received a lot o…

regression

Understanding the Mixture-of-Experts with Nadaraya-Watson Kernel

2025-09-30 · Chuanyang Zheng, Jiankai Sun, Yihang Gao, Enze Xie 외 arxiv

Mixture-of-Experts (MoE) has become a cornerstone in recent state-of-the-art large language models (LLMs). Traditionally, MoE relies on $\mathrm{Softmax}$ as the router score function to aggregate expert output, a design…