paper-with-me

홈 › Papers

Optimal Ridge Regularization for Out-of-Distribution Prediction

2024-04-01 · Pratik Patil, Jin-Hong Du, Ryan J. Tibshirani

We study the behavior of optimal ridge regularization and optimal ridge risk for out-of-distribution prediction, where the test distribution deviates arbitrarily from the train distribution. We establish general conditions that determine the sign of the optimal regularization level under covariate and regression shifts. These conditions capture the alignment between the covariance and signal structures in the train and test data and reveal stark differences compared to the in-distribution setting. For example, a negative regularization level can be optimal under covariate shift or regression shift, even when the training features are isotropic or the design is underparameterized. Furthermore, we prove that the optimally-tuned risk is monotonic in the data aspect ratio, even in the out-of-distribution setting and when optimizing over negative regularization levels. In general, our results do not make any modeling assumptions for the train or the test distributions, except for moment bounds, and allow for arbitrary shifts and the widest possible range of (negative) regularization levels.

📄 PDF Abstract BibTeX arXiv:2404.01233

Code (1)

jaydu1/ood-ridge 공식 구현

Tasks

Predictionregression

Similar Papers 제목 키워드 기반

Generalized equivalences between subsampling and ridge regularization

2023-05-29 · NeurIPS 2023 11

We establish precise structural and risk equivalences between subsampling and ridge regularization for ensemble ridge estimators. Specifically, we prove that linear and quadratic functionals of subsample ridge estimators…

regression

Optimal ridge regularization revisited

2026-05-27 · Jack Timmermans, Sergio A. Alvarez arxiv

We consider $L^2$-regularized linear (ridge) regression over a finite data sample $X$ with bounded covariance and linear prediction targets $y$ with additive isotropic noise of finite variance. We present an iterative pr…

Fundamental Limits of Ridge-Regularized Empirical Risk Minimization in High Dimensions

2020-06-16 · Hossein Taheri, Ramtin Pedarsani, Christos Thrampoulidis

Empirical Risk Minimization (ERM) algorithms are widely used in a variety of estimation and prediction tasks in signal-processing and machine learning applications. Despite their popularity, a theory that explains their …

Vocal Bursts Intensity Prediction

The Ridge Path Estimator for Linear Instrumental Variables

2019-08-25 · Nandana Sengupta, Fallaw Sowell

This paper presents the asymptotic behavior of a linear instrumental variables (IV) estimator that uses a ridge regression penalty. The regularization tuning parameter is selected empirically by splitting the observed da…

Source-Optimal Training is Transfer-Suboptimal

2025-11-11 · C. Evans Hedges arxiv

We prove that training a source model optimally for its own task is generically suboptimal when the objective is downstream transfer. We study the source-side optimization problem in L2-SP ridge regression and show a fun…

Transfer Learning