paper-with-me

Papers

Finite-sample Analysis of Interpolating Linear Classifiers in the Overparameterized Regime

2020-04-25 · Niladri S. Chatterji, Philip M. Long

We prove bounds on the population risk of the maximum margin algorithm for two-class linear classification. For linearly separable training data, the maximum margin algorithm has been shown in previous work to be equivalent to a limit of training with logistic loss using gradient descent, as the training error is driven to zero. We analyze this algorithm applied to random data including misclassification noise. Our assumptions on the clean data include the case in which the class-conditional distributions are standard normal distributions. The misclassification noise may be chosen by an adversary, subject to a limit on the fraction of corrupted labels. Our bounds show that, with sufficient over-parameterization, the maximum margin algorithm trained on noisy data can achieve nearly optimal population risk.

📄 PDF Abstract BibTeX arXiv:2004.12019

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Interpolating Classifiers Make Few Mistakes

2021-01-28 · Tengyuan Liang, Benjamin Recht

This paper provides elementary analyses of the regret and generalization of minimum-norm interpolating classifiers (MNIC). The MNIC is the function of smallest Reproducing Kernel Hilbert Space norm that perfectly interpo…

Infinite Class Mixup

2023-05-17 · Thomas Mensink, Pascal Mettes

Mixup is a widely adopted strategy for training deep networks, where additional samples are augmented by interpolating inputs and labels of training pairs. Mixup has shown to improve classification performance, network c…

Fast Convergence in Learning Two-Layer Neural Networks with Separable Data

2023-05-22 · Hossein Taheri, Christos Thrampoulidis

Normalized gradient descent has shown substantial success in speeding up the convergence of exponentially-tailed loss functions (which includes exponential and logistic losses) on linear classifiers with separable data. …

Generalization Bounds

Interpolating Predictors in High-Dimensional Factor Regression

2020-02-06 · Florentina Bunea, Seth Strimas-Mackey, Marten Wegkamp

This work studies finite-sample properties of the risk of the minimum-norm interpolating predictor in high-dimensional regression models. If the effective rank of the covariance matrix $\Sigma$ of the $p$ regression feat…

regressionVocal Bursts Intensity Prediction

Benign Overfitting in Linear Regression

2019-06-26 · Peter L. Bartlett, Philip M. Long, Gábor Lugosi, Alexander Tsigler

The phenomenon of benign overfitting is one of the key mysteries uncovered by deep learning methodology: deep neural networks seem to predict well, even with a perfect fit to noisy training data. Motivated by this phenom…

Predictionregression