paper-with-me

홈 › Papers

Selecting the number of clusters, clustering models, and algorithms. A unifying approach based on the quadratic discriminant score

2021-11-03 · Luca Coraggio, Pietro Coretto

Cluster analysis requires many decisions: the clustering method and the implied reference model, the number of clusters and, often, several hyper-parameters and algorithms' tunings. In practice, one produces several partitions, and a final one is chosen based on validation or selection criteria. There exist an abundance of validation methods that, implicitly or explicitly, assume a certain clustering notion. Moreover, they are often restricted to operate on partitions obtained from a specific method. In this paper, we focus on groups that can be well separated by quadratic or linear boundaries. The reference cluster concept is defined through the quadratic discriminant score function and parameters describing clusters' size, center and scatter. We develop two cluster-quality criteria called quadratic scores. We show that these criteria are consistent with groups generated from a general class of elliptically-symmetric distributions. The quest for this type of groups is common in applications. The connection with likelihood theory for mixture models and model-based clustering is investigated. Based on bootstrap resampling of the quadratic scores, we propose a selection rule that allows choosing among many clustering solutions. The proposed method has the distinctive advantage that it can compare partitions that cannot be compared with other state-of-the-art methods. Extensive numerical experiments and the analysis of real data show that, even if some competing methods turn out to be superior in some setups, the proposed methodology achieves a better overall performance.

📄 PDF Abstract BibTeX arXiv:2111.02302

Code (0)

등록된 구현이 없습니다.

Tasks

Clustering

Similar Papers 제목 키워드 기반

Review: Metaheuristic Search-Based Fuzzy Clustering Algorithms

2018-01-21 · Waleed Alomoush, Ayat Alrosan

Fuzzy clustering is a famous unsupervised learning method used to collecting similar data elements within cluster according to some similarity measurement. But, clustering algorithms suffer from some drawbacks. Among the…

Clustering

Dynamic Clustering in Federated Learning

2020-12-07 · Yeongwoo Kim, Ezeddin Al Hakim, Johan Haraldson, Henrik Eriksson 외

In the resource management of wireless networks, Federated Learning has been used to predict handovers. However, non-independent and identically distributed data degrade the accuracy performance of such predictions. To o…

ClusteringFederated LearningGenerative Adversarial NetworkManagement+3

Clustering by Sum of Norms: Stochastic Incremental Algorithm, Convergence and Cluster Recovery

2017-08-01 · ICML 2017 8 · Ashkan Panahi, Devdatt Dubhashi, Fredrik D. Johansson, Chiranjib Bhattacharyya

Standard clustering methods such as K-means, Gaussian mixture models, and hierarchical clustering are beset by local minima, which are sometimes drastically suboptimal. Moreover the number of clusters K must be know…

Clustering

Clustering Algorithms to Analyze the Road Traffic Crashes

2021-08-07 · Mahnaz Rafia Islam, Israt Jahan Jenny, Moniruzzaman Nayon, Md. Rajibul Islam 외

Selecting an appropriate clustering method as well as an optimal number of clusters in road accident data is at times confusing and difficult. This paper analyzes shortcomings of different existing techniques applied to …

Clustering

Clustering Aggregation as Maximum-Weight Independent Set

2012-12-01 · NeurIPS 2012 12 · Nan Li, Longin J. Latecki

We formulate clustering aggregation as a special instance of Maximum-Weight Independent Set (MWIS) problem. For a given dataset, an attributed graph is constructed from the union of the input clusterings generated by dif…

Clustering