paper-with-me

Papers

Sparse Probit Linear Mixed Model

2015-07-16 · Stephan Mandt, Florian Wenzel, Shinichi Nakajima, John P. Cunningham, Christoph Lippert, Marius Kloft

Linear Mixed Models (LMMs) are important tools in statistical genetics. When used for feature selection, they allow to find a sparse set of genetic traits that best predict a continuous phenotype of interest, while simultaneously correcting for various confounding factors such as age, ethnicity and population structure. Formulated as models for linear regression, LMMs have been restricted to continuous phenotypes. We introduce the Sparse Probit Linear Mixed Model (Probit-LMM), where we generalize the LMM modeling paradigm to binary phenotypes. As a technical challenge, the model no longer possesses a closed-form likelihood function. In this paper, we present a scalable approximate inference algorithm that lets us fit the model to high-dimensional data sets. We show on three real-world examples from different domains that in the setup of binary labels, our algorithm leads to better prediction accuracies and also selects features which show less correlation with the confounding factors.

📄 PDF Abstract BibTeX arXiv:1507.04777

Code (0)

등록된 구현이 없습니다.

Tasks

feature selectionmodel

Similar Papers 제목 키워드 기반

Learning sparse generalized linear models with binary outcomes via iterative hard thresholding

2025-02-25 · Namiko Matsumoto, Arya Mazumdar

In statistics, generalized linear models (GLMs) are widely used for modeling data and can expressively capture potential nonlinear dependence of the model's outcomes on its covariates. Within the broad family of GLMs, th…

Binary Classificationparameter estimationregression

Linearized Binary Regression

2018-02-01 · Andrew S. Lan, Mung Chiang, Christoph Studer

Probit regression was first proposed by Bliss in 1934 to study mortality rates of insects. Since then, an extensive body of work has analyzed and used probit or related binary regression methods (such as logistic regress…

regression

Generalization Bounds and Consistency for Latent Structural Probit and Ramp Loss

2011-12-01 · NeurIPS 2011 12 · Joseph Keshet, David A. Mcallester

We consider latent structural versions of probit loss and ramp loss. We show that these surrogate loss functions are consistent in the strong sense that for any feature map (finite or infinite dimensional) they yield p…

Generalization Bounds

Robust Ranking of Happiness Outcomes: A Median Regression Perspective

2019-02-20 · Le-Yu Chen, Ekaterina Oparina, Nattavudh Powdthavee, Sorawoot Srisuma

Ordered probit and logit models have been frequently used to estimate the mean ranking of happiness outcomes (and other ordinal data) across groups. However, it has been recently highlighted that such ranking may not be …

regression

$p$-Generalized Probit Regression and Scalable Maximum Likelihood Estimation via Sketching and Coresets

2022-03-25 · Alexander Munteanu, Simon Omlor, Christian Peters

We study the $p$-generalized probit regression model, which is a generalized linear model for binary responses. It extends the standard probit model by replacing its link function, the standard normal cdf, by a $p$-gener…

regression