paper-with-me

Papers

Interpolating Predictors in High-Dimensional Factor Regression

2020-02-06 · Florentina Bunea, Seth Strimas-Mackey, Marten Wegkamp

This work studies finite-sample properties of the risk of the minimum-norm interpolating predictor in high-dimensional regression models. If the effective rank of the covariance matrix $\Sigma$ of the $p$ regression features is much larger than the sample size $n$, we show that the min-norm interpolating predictor is not desirable, as its risk approaches the risk of trivially predicting the response by 0. However, our detailed finite-sample analysis reveals, surprisingly, that this behavior is not present when the regression response and the features are {\it jointly} low-dimensional, following a widely used factor regression model. Within this popular model class, and when the effective rank of $\Sigma$ is smaller than $n$, while still allowing for $p \gg n$, both the bias and the variance terms of the excess risk can be controlled, and the risk of the minimum-norm interpolating predictor approaches optimal benchmarks. Moreover, through a detailed analysis of the bias term, we exhibit model classes under which our upper bound on the excess risk approaches zero, while the corresponding upper bound in the recent work arXiv:1906.11300 diverges. Furthermore, we show that the minimum-norm interpolating predictor analyzed under the factor regression model, despite being model-agnostic and devoid of tuning parameters, can have similar risk to predictors based on principal components regression and ridge regression, and can improve over LASSO based predictors, in the high-dimensional regime.

📄 PDF Abstract BibTeX arXiv:2002.02525

Code (0)

등록된 구현이 없습니다.

Tasks

regressionVocal Bursts Intensity Prediction

Similar Papers 제목 키워드 기반

Prediction in latent factor regression: Adaptive PCR and beyond

2020-07-20 · Xin Bing, Florentina Bunea, Seth Strimas-Mackey, Marten Wegkamp

This work is devoted to the finite sample prediction risk analysis of a class of linear predictors of a response $Y\in \mathbb{R}$ from a high-dimensional random vector $X\in \mathbb{R}^p$ when $(X,Y)$ follows a latent f…

Model SelectionPredictionregression

Kernel Three Pass Regression Filter

2024-05-12 · Rajveer Jat, Daanish Padha

We forecast a single time series using a high-dimensional set of predictors. When these predictors share common underlying dynamics, an approximate latent factor model provides a powerful characterization of their co-mov…

regression

Sufficient Forecasting Using Factor Models

2015-05-27 · Jianqing Fan, Lingzhou Xue, Jiawei Yao

We consider forecasting a single time series when there is a large number of predictors and a possible nonlinear effect. The dimensionality was first reduced via a high-dimensional (approximate) factor model implemented …

Dimensionality ReductionregressionTime Series Analysis

A Supervised Screening and Regularized Factor-Based Method for Time Series Forecasting

2025-02-21 · Sihan Tu, Zhaoxing Gao

Factor-based forecasting using Principal Component Analysis (PCA) is an effective machine learning tool for dimension reduction with many applications in statistics, economics, and finance. This paper introduces a Superv…

Dimensionality ReductionTime SeriesTime Series ForecastingTime Series Regression

On Uniform Convergence and Low-Norm Interpolation Learning

2020-06-10 · NeurIPS 2020 12 · Lijia Zhou, Danica J. Sutherland, Nathan Srebro

We consider an underdetermined noisy linear regression model where the minimum-norm interpolating predictor is known to be consistent, and ask: can uniform convergence in a norm ball, or at least (following Nagarajan and…