paper-with-me

Papers

Robust Information Criterion for Model Selection in Sparse High-Dimensional Linear Regression Models

2022-06-17 · Prakash B. Gohain, Magnus Jansson

Model selection in linear regression models is a major challenge when dealing with high-dimensional data where the number of available measurements (sample size) is much smaller than the dimension of the parameter space. Traditional methods for model selection such as Akaike information criterion, Bayesian information criterion (BIC) and minimum description length are heavily prone to overfitting in the high-dimensional setting. In this regard, extended BIC (EBIC), which is an extended version of the original BIC and extended Fisher information criterion (EFIC), which is a combination of EBIC and Fisher information criterion, are consistent estimators of the true model as the number of measurements grows very large. However, EBIC is not consistent in high signal-to-noise-ratio (SNR) scenarios where the sample size is fixed and EFIC is not invariant to data scaling resulting in unstable behaviour. In this paper, we propose a new form of the EBIC criterion called EBIC-Robust, which is invariant to data scaling and consistent in both large sample size and high-SNR scenarios. Analytical proofs are presented to guarantee its consistency. Simulation results indicate that the performance of EBIC-Robust is quite superior to that of both EBIC and EFIC.

📄 PDF Abstract BibTeX arXiv:2206.08731

Code (0)

등록된 구현이 없습니다.

Tasks

Model Selectionregression

Methods 이 논문이 사용한 방법론

Linear Regression Linear Regression is a method for modelling a relationship between a dependent variable and independent variables. These models can be fit with numerous approaches. The most…

Similar Papers 제목 키워드 기반

Model Selection in High-Dimensional Block-Sparse Linear Regression

2022-09-03 · Prakash B. Gohain, Magnus Jansson

Model selection is an indispensable part of data analysis dealing very frequently with fitting and prediction purposes. In this paper, we tackle the problem of model selection in a general linear regression where the par…

Model SelectionregressionVocal Bursts Intensity Prediction

Improving Group Lasso for high-dimensional categorical data

2022-10-25 · Szymon Nowakowski, Piotr Pokarowski, Wojciech Rejchel, Agnieszka Sołtys

Sparse modelling or model selection with categorical data is challenging even for a moderate number of variables, because one parameter is roughly needed to encode one category or level. The Group Lasso is a well known e…

Model SelectionVocal Bursts Intensity Prediction

Quick and Robust Feature Selection: the Strength of Energy-efficient Sparse Training for Autoencoders

2020-12-01 · Zahra Atashgahi, Ghada Sokar, Tim Van der Lee, Elena Mocanu 외

Major complications arise from the recent increase in the amount of high-dimensional data, including high computational costs and memory requirements. Feature selection, which identifies the most relevant and informative…

ClusteringDenoisingDimensionality ReductionFeature Importance+1

Stability Approach to Regularization Selection (StARS) for High Dimensional Graphical Models

2010-06-16 · NeurIPS 2010 12 · Han Liu, Kathryn Roeder, Larry Wasserman

A challenging problem in estimating high-dimensional graphical models is to choose the regularization parameter in a data-dependent way. The standard techniques include $K$-fold cross-validation ($K$-CV), Akaike informat…

Model SelectionVocal Bursts Intensity Prediction

Model selection for high-dimensional linear regression with dependent observations

2019-06-18 · Ching-Kang Ing

We investigate the prediction capability of the orthogonal greedy algorithm (OGA) in high-dimensional regression models with dependent observations. The rates of convergence of the prediction error of OGA are obtained un…

Model SelectionPredictionregressionVocal Bursts Intensity Prediction