paper-with-me

홈 › Papers

Survey of state-of-the-art mixed data clustering algorithms

2018-11-11 · Amir Ahmad, Shehroz S. Khan

Mixed data comprises both numeric and categorical features, and mixed datasets occur frequently in many domains, such as health, finance, and marketing. Clustering is often applied to mixed datasets to find structures and to group similar objects for further analysis. However, clustering mixed data is challenging because it is difficult to directly apply mathematical operations, such as summation or averaging, to the feature values of these datasets. In this paper, we present a taxonomy for the study of mixed data clustering algorithms by identifying five major research themes. We then present a state-of-the-art review of the research works within each research theme. We analyze the strengths and weaknesses of these methods with pointers for future research directions. Lastly, we present an in-depth analysis of the overall challenges in this field, highlight open research questions and discuss guidelines to make progress in the field.

📄 PDF Abstract BibTeX arXiv:1811.04364

Code (0)

등록된 구현이 없습니다.

Tasks

ClusteringMarketingSurvey

Similar Papers 제목 키워드 기반

Mixed Data Clustering Survey and Challenges

2025-11-27 · Guillaume Guerard, Sonia Djebali arxiv

The advent of the big data paradigm has transformed how industries manage and analyze information, ushering in an era of unprecedented data volume, velocity, and variety. Within this landscape, mixed-data clustering has …

initKmix -- A Novel Initial Partition Generation Algorithm for Clustering Mixed Data using k-means-based Clustering

2019-01-31 · Amir Ahmad, Shehroz S. Khan

Mixed datasets consist of both numeric and categorical attributes. Various k-means-based clustering algorithms have been developed for these datasets. Generally, these algorithms use random partition as a starting point,…

Clustering

Hybrid Density- and Partition-based Clustering Algorithm for Data with Mixed-type Variables

2019-05-06 · Shu Wang, Jonathan G. Yabes, Chung-Chou H. Chang

Clustering is an essential technique for discovering patterns in data. The steady increase in amount and complexity of data over the years led to improvements and development of new clustering algorithms. However, algori…

Clustering

A Survey on Soft Subspace Clustering

2014-09-19 · Zhaohong Deng, Kup-Sze Choi, Yizhang Jiang, Jun Wang 외

Subspace clustering (SC) is a promising clustering technology to identify clusters based on their associations with subspaces in high dimensional spaces. SC can be classified into hard subspace clustering (HSC) and soft …

ClusteringSurvey

Unsupervised Diffusion and Volume Maximization-Based Clustering of Hyperspectral Images

2022-03-18 · Sam L. Polk, Kangning Cui, Aland H. Y. Chan, David A. Coomes 외

Hyperspectral images taken from aircraft or satellites contain information from hundreds of spectral bands, within which lie latent lower-dimensional structures that can be exploited for classifying vegetation and other …

ClusteringImage Clustering