paper-with-me

홈 › Papers

Optimal and Safe Estimation for High-Dimensional Semi-Supervised Learning

2020-11-28 · Siyi Deng, Yang Ning, Jiwei Zhao, Heping Zhang

We consider the estimation problem in high-dimensional semi-supervised learning. Our goal is to investigate when and how the unlabeled data can be exploited to improve the estimation of the regression parameters of linear model in light of the fact that such linear models may be misspecified in data analysis. We first establish the minimax lower bound for parameter estimation in the semi-supervised setting, and show that this lower bound cannot be achieved by supervised estimators using the labeled data only. We propose an optimal semi-supervised estimator that can attain this lower bound and therefore improves the supervised estimators, provided that the conditional mean function can be consistently estimated with a proper rate. We further propose a safe semi-supervised estimator. We view it safe, because this estimator is always at least as good as the supervised estimators. We also extend our idea to the aggregation of multiple semi-supervised estimators caused by different misspecifications of the conditional mean function. Extensive numerical simulations and a real data analysis are conducted to illustrate our theoretical results.

📄 PDF Abstract BibTeX arXiv:2011.14185

Code (0)

등록된 구현이 없습니다.

Tasks

parameter estimationregressionVocal Bursts Intensity Prediction

Methods 이 논문이 사용한 방법론

Linear Regression Linear Regression is a method for modelling a relationship between a dependent variable and independent variables. These models can be fit with numerous approaches. The most…

Similar Papers 제목 키워드 기반

Semi-Supervised Quantile Estimation: Robust and Efficient Inference in High Dimensional Settings

2022-01-25 · Abhishek Chakrabortty, Guorong Dai, Raymond J. Carroll

We consider quantile estimation in a semi-supervised setting, characterized by two available data sets: (i) a small or moderate sized labeled data set containing observations for a response and a set of possibly high dim…

Dimensionality ReductionImputation

High-Dimensional Variance-Reduced Stochastic Gradient Expectation-Maximization Algorithm

2017-08-01 · ICML 2017 8 · Rongda Zhu, Lingxiao Wang, ChengXiang Zhai, Quanquan Gu

We propose a generic stochastic expectation-maximization (EM) algorithm for the estimation of high-dimensional latent variable models. At the core of our algorithm is a novel semi-stochastic variance-reduced gradien…

parameter estimationVocal Bursts Intensity Prediction

High-dimensional semi-supervised learning: in search for optimal inference of the mean

2019-02-02 · Yuqian Zhang, Jelena Bradic

We provide a high-dimensional semi-supervised inference framework focused on the mean and variance of the response. Our data are comprised of an extensive set of observations regarding the covariate vectors and a much sm…

High Dimensional M-Estimation with Missing Outcomes: A Semi-Parametric Framework

2019-11-26 · Abhishek Chakrabortty, Jiarui Lu, T. Tony Cai, Hongzhe Li

We consider high dimensional $M$-estimation in settings where the response $Y$ is possibly missing at random and the covariates $\mathbf{X} \in \mathbb{R}^p$ can be high dimensional compared to the sample size $n$. The p…

Causal InferenceregressionVocal Bursts Intensity Prediction

DNA-SE: Towards Deep Neural-Nets Assisted Semiparametric Estimation

2024-08-04 · Qinshuo Liu, Zixin Wang, Xi-An Li, Xinyao Ji 외

Semiparametric statistics play a pivotal role in a wide range of domains, including but not limited to missing data, causal inference, and transfer learning, to name a few. In many settings, semiparametric theory leads t…

Causal InferenceTransfer Learning