paper-with-me

홈 › Papers

Sparse High-Dimensional Regression: Exact Scalable Algorithms and Phase Transitions

2017-09-28 · Dimitris Bertsimas, Bart Van Parys

We present a novel binary convex reformulation of the sparse regression problem that constitutes a new duality perspective. We devise a new cutting plane method and provide evidence that it can solve to provable optimality the sparse regression problem for sample sizes n and number of regressors p in the 100,000s, that is two orders of magnitude better than the current state of the art, in seconds. The ability to solve the problem for very high dimensions allows us to observe new phase transition phenomena. Contrary to traditional complexity theory which suggests that the difficulty of a problem increases with problem size, the sparse regression problem has the property that as the number of samples $n$ increases the problem becomes easier in that the solution recovers 100% of the true signal, and our approach solves the problem extremely fast (in fact faster than Lasso), while for small number of samples n, our approach takes a larger amount of time to solve the problem, but importantly the optimal solution provides a statistically more relevant regressor. We argue that our exact sparse regression approach presents a superior alternative over heuristic methods available at present.

📄 PDF Abstract BibTeX arXiv:1709.10029

Code (0)

등록된 구현이 없습니다.

Tasks

regressionVocal Bursts Intensity Prediction

Similar Papers 제목 키워드 기반

Kernel Packet: An Exact and Scalable Algorithm for Gaussian Process Regression with Matérn Correlations

2022-03-07 · HaoYuan Chen, Liang Ding, Rui Tuo

We develop an exact and scalable algorithm for one-dimensional Gaussian process regression with Mat\'ern correlations whose smoothness parameter $\nu$ is a half-integer. The proposed algorithm only requires $\mathcal{O}(…

regression

Scalable Matrix-valued Kernel Learning for High-dimensional Nonlinear Multivariate Regression and Granger Causality

2014-08-09 · Vikas Sindhwani, Ha Quang Minh, Aurelie Lozano

We propose a general matrix-valued multiple kernel learning framework for high-dimensional nonlinear multivariate regression problems. This framework allows a broad class of mixed norm regularizers, including those that …

Causal InferenceGeneralization Boundsregression

A Consistent and Scalable Algorithm for Best Subset Selection in Single Index Models

2023-09-12 · Borui Tang, Jin Zhu, Junxian Zhu, Xueqin Wang 외

Analysis of high-dimensional data has led to increased interest in both single index models (SIMs) and best subset selection. SIMs provide an interpretable and flexible modeling framework for high-dimensional data, while…

Model Selectionregression

Inference in high-dimensional regression models without the exact or $L^p$ sparsity

2021-08-21 · Jooyoung Cha, Harold D. Chiang, Yuya Sasaki

This paper proposes a new method of inference in high-dimensional regression models and high-dimensional IV regression models. Estimation is based on a combined use of the orthogonal greedy algorithm, high-dimensional Ak…

regression

TabNSM: Neural Sparse Mixer for Tabular Regression

2026-08-18 · Ali Eslamian, Qiang Cheng arxiv

Large-scale, high-dimensional tabular regression remains challenging: tree-based models are robust but lack end-to-end representation learning, while deep models enable flexible feature learning but often incur costly in…

Representation Learning