Estimation and Feature Selection in Mixtures of Generalized Linear Experts Models
Mixtures-of-Experts (MoE) are conditional mixture models that have shown their performance in modeling heterogeneity in data in many statistical learning approaches for prediction, including regression and classification, as well as for clustering. Their estimation in high-dimensional problems is still however challenging. We consider the problem of parameter estimation and feature selection in MoE models with different generalized linear experts models, and propose a regularized maximum likelihood estimation that efficiently encourages sparse solutions for heterogeneous data with high-dimensional predictors. The developed proximal-Newton EM algorithm includes proximal Newton-type procedures to update the model parameter by monotonically maximizing the objective function and allows to perform efficient estimation and feature selection. An experimental study shows the good performance of the algorithms in terms of recovering the actual sparse solutions, parameter estimation, and clustering of heterogeneous regression data, compared to the main state-of-the art competitors.
Code (1)
Tasks
Clusteringfeature selectionparameter estimationregressionMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Fairness Constraints in High-Dimensional Generalized Linear Models
Most fairness-aware learning methods assume that sensitive attributes are observed, an assumption that may fail due to privacy, legal, or data-collection constraints. We develop a framework for fairness-aware generalized…
Least Angle Regression in Tangent Space and LASSO for Generalized Linear Models
This study proposes sparse estimation methods for the generalized linear models, which run one of least angle regression (LARS) and least absolute shrinkage and selection operator (LASSO) in the tangent space of the mani…
Model Selectionparameter estimationregressionVariable SelectionProvable Tensor Methods for Learning Mixtures of Generalized Linear Models
We consider the problem of learning mixtures of generalized linear models (GLM) which arise in classification and regression problems. Typical learning approaches such as expectation maximization (EM) or variational Baye…
General ClassificationTensor DecompositionDensity Ratio Estimation via Sampling along Generalized Geodesics on Statistical Manifolds
The density ratio of two probability distributions is one of the fundamental tools in mathematical and computational statistics and machine learning, and it has a variety of known applications. Therefore, density ratio e…
Density Ratio EstimationAutomated Model Selection for Generalized Linear Models
In this paper, we show how mixed-integer conic optimization can be used to combine feature subset selection with holistic generalized linear models to fully automate the model selection process. Concretely, we directly o…
feature selectionmodelModel Selectionregression