paper-with-me

Papers

Mixture Model Averaging for Clustering

2012-12-23 · Yuhong Wei, Paul D. McNicholas

In mixture model-based clustering applications, it is common to fit several models from a family and report clustering results from only the `best' one. In such circumstances, selection of this best model is achieved using a model selection criterion, most often the Bayesian information criterion. Rather than throw away all but the best model, we average multiple models that are in some sense close to the best one, thereby producing a weighted average of clustering results. Two (weighted) averaging approaches are considered: averaging the component membership probabilities and averaging models. In both cases, Occam's window is used to determine closeness to the best model and weights are computed within a Bayesian model averaging paradigm. In some cases, we need to merge components before averaging; we introduce a method for merging mixture components based on the adjusted Rand index. The effectiveness of our model-based clustering averaging approaches is illustrated using a family of Gaussian mixture models on real and simulated data.

📄 PDF Abstract BibTeX arXiv:1212.5760

Code (0)

등록된 구현이 없습니다.

Tasks

ClusteringmodelModel Selection

Similar Papers 제목 키워드 기반

Strength in Numbers: Averaging and Clustering Effects in Mixture of Experts for Graph-Based Dependency Parsing

2021-08-01 · ACL (IWPT) 2021 8 · Xudong Zhang, Joseph Le Roux, Thierry Charnois

We review two features of mixture of experts (MoE) models which we call averaging and clustering effects in the context of graph-based dependency parsers learned in a supervised probabilistic framework. Averaging corresp…

ClusteringDependency ParsingMixture-of-Experts

BET: Bayesian Ensemble Trees for Clustering and Prediction in Heterogeneous Data

2014-08-18 · Leo L. Duan, John P. Clancy, Rhonda D. Szczesniak

We propose a novel "tree-averaging" model that utilizes the ensemble of classification and regression trees (CART). Each constituent tree is estimated with a subset of similar data. We treat this grouping of subsets as B…

ClassificationClusteringGeneral Classificationregression

Honeyfile Camouflage: Hiding Fake Files in Plain Sight

2024-05-08 · Roelien C. Timmer, David Liebowitz, Surya Nepal, Salil S. Kanhere

Honeyfiles are a particularly useful type of honeypot: fake files deployed to detect and infer information from malicious behaviour. This paper considers the challenge of naming honeyfiles so they are camouflaged when pl…

Clustering

Mixture model modal clustering

2016-09-15 · José E. Chacón

The two most extended density-based approaches to clustering are surely mixture model clustering and modal clustering. In the mixture model approach, the density is represented as a mixture and clusters are associated to…

Clusteringmodel

Comparison of Clustering Algorithms for Statistical Features of Vibration Data Sets

2023-05-11 · Philipp Sepin, Jana Kemnitz, Safoura Rezapour Lakani, Daniel Schall

Vibration-based condition monitoring systems are receiving increasing attention due to their ability to accurately identify different conditions by capturing dynamic features over a broad frequency range. However, there …

Clusteringfeature selection