paper-with-me

홈 › Papers

Harmless interpolation in regression and classification with structured features

2021-11-09 · Andrew D. McRae, Santhosh Karnik, Mark A. Davenport, Vidya Muthukumar

Overparametrized neural networks tend to perfectly fit noisy training data yet generalize well on test data. Inspired by this empirical observation, recent work has sought to understand this phenomenon of benign overfitting or harmless interpolation in the much simpler linear model. Previous theoretical work critically assumes that either the data features are statistically independent or the input data is high-dimensional; this precludes general nonparametric settings with structured feature maps. In this paper, we present a general and flexible framework for upper bounding regression and classification risk in a reproducing kernel Hilbert space. A key contribution is that our framework describes precise sufficient conditions on the data Gram matrix under which harmless interpolation occurs. Our results recover prior independent-features results (with a much simpler analysis), but they furthermore show that harmless interpolation can occur in more general settings such as features that are a bounded orthonormal system. Furthermore, our results show an asymptotic separation between classification and regression performance in a manner that was previously only shown for Gaussian features.

📄 PDF Abstract BibTeX arXiv:2111.05198

Code (0)

등록된 구현이 없습니다.

Tasks

Classificationregression

Similar Papers 제목 키워드 기반

Strong inductive biases provably prevent harmless interpolation

2023-01-18 · Michael Aerni, Marco Milanta, Konstantin Donhauser, Fanny Yang

Classical wisdom suggests that estimators should avoid fitting noise to achieve good generalization. In contrast, modern overparameterized models can yield small test error despite interpolating noise -- a phenomenon oft…

Inductive Bias

Harmless interpolation of noisy data in regression

2019-03-21 · Vidya Muthukumar, Kailas Vodrahalli, Vignesh Subramanian, Anant Sahai

A continuing mystery in understanding the empirical success of deep neural networks is their ability to achieve zero training error and generalize well, even when the training data is noisy and there are more parameters …

regression

New Equivalences Between Interpolation and SVMs: Kernels and Structured Features

2023-05-03 · Chiraag Kaushik, Andrew D. McRae, Mark A. Davenport, Vidya Muthukumar

The support vector machine (SVM) is a supervised learning algorithm that finds a maximum-margin linear classifier, often after mapping the data to a high-dimensional feature space via the kernel trick. Recent work has de…

Merging Two Cultures: Deep and Statistical Learning

2021-10-22 · Anindya Bhadra, Jyotishka Datta, Nick Polson, Vadim Sokolov 외

Merging the two cultures of deep and statistical learning provides insights into structured high-dimensional data. Traditional statistical modeling is still a dominant strategy for structured tabular data. Deep learning …

Dimensionality ReductionFeature EngineeringUncertainty QuantificationVocal Bursts Valence Prediction

Task Shift: From Classification to Regression in Overparameterized Linear Models

2025-02-18 · Tyler LaBonte, Kuo-Wei Lai, Vidya Muthukumar

Modern machine learning methods have recently demonstrated remarkable capability to generalize under task shift, where latent knowledge is transferred to a different, often more difficult, task under a similar data distr…

regression