paper-with-me

홈 › Papers

Non-asymptotic model selection in block-diagonal mixture of polynomial experts models

2021-04-18 · TrungTin Nguyen, Faicel Chamroukhi, Hien Duy Nguyen, Florence Forbes

Model selection, via penalized likelihood type criteria, is a standard task in many statistical inference and machine learning problems. Progress has led to deriving criteria with asymptotic consistency results and an increasing emphasis on introducing non-asymptotic criteria. We focus on the problem of modeling non-linear relationships in regression data with potential hidden graph-structured interactions between the high-dimensional predictors, within the mixture of experts modeling framework. In order to deal with such a complex situation, we investigate a block-diagonal localized mixture of polynomial experts (BLoMPE) regression model, which is constructed upon an inverse regression and block-diagonal structures of the Gaussian expert covariance matrices. We introduce a penalized maximum likelihood selection criterion to estimate the unknown conditional density of the regression model. This model selection criterion allows us to handle the challenging problem of inferring the number of mixture components, the degree of polynomial mean functions, and the hidden block-diagonal structures of the covariance matrices, which reduces the number of parameters to be estimated and leads to a trade-off between complexity and sparsity in the model. In particular, we provide a strong theoretical guarantee: a finite-sample oracle inequality satisfied by the penalized maximum likelihood estimator with a Jensen-Kullback-Leibler type loss, to support the introduced non-asymptotic model selection criterion. The penalty shape of this criterion depends on the complexity of the considered random subcollection of BLoMPE models, including the relevant graph structures, the degree of polynomial mean functions, and the number of mixture components.

📄 PDF Abstract BibTeX arXiv:2104.08959

Code (0)

등록된 구현이 없습니다.

Tasks

Mixture-of-ExpertsModel Selectionregression

Similar Papers 제목 키워드 기반

A non-asymptotic approach for model selection via penalization in high-dimensional mixture of experts models

2021-04-06 · TrungTin Nguyen, Hien Duy Nguyen, Faicel Chamroukhi, Florence Forbes

Mixture of experts (MoE) are a popular class of statistical and machine learning models that have gained attention over the years due to their flexibility and efficiency. In this work, we consider Gaussian-gated localize…

Mixture-of-ExpertsModel Selection

A Comparison of Variable Selection Methods for Blockwise Diagonal Designs

2021-09-29 · ICLR 2022 4 · Tracy Ke, Longlin Wang

Lasso is a celebrated method for variable selection in linear models, but it faces challenges when the covariates are moderately or strongly correlated. This motivates alternative approaches such as using a non-convex pe…

Variable Selection

Block-diagonal covariance selection for high-dimensional Gaussian graphical models

2015-11-12 · Emilie Devijver, Mélina Gallopin

Gaussian graphical models are widely utilized to infer and visualize networks of dependencies between continuous variables. However, inferring the graph is difficult when the sample size is small compared to the number o…

Dimensionality ReductionModel SelectionVocal Bursts Intensity Prediction

Phase Transitions and a Model Order Selection Criterion for Spectral Graph Clustering

2016-04-11 · Pin-Yu Chen, Alfred O. Hero

One of the longstanding open problems in spectral graph clustering (SGC) is the so-called model order selection problem: automated selection of the correct number of clusters. This is equivalent to the problem of finding…

ClusteringGraph ClusteringModel SelectionSpectral Graph Clustering

Gaussian Mixture Model with unknown diagonal covariances via continuous sparse regularization

2025-09-16 · Romane Giard, Yohann de Castro, Clément Marteau arxiv

This paper addresses the statistical estimation of Gaussian Mixture Models (GMMs) with unknown diagonal covariances from independent and identically distributed samples. We employ the Beurling-LASSO (BLASSO), a convex op…