paper-with-me

홈 › Papers

Feature selection or extraction decision process for clustering using PCA and FRSD

2021-11-20 · Jean-Sebastien Dessureault, Daniel Massicotte

This paper concerns the critical decision process of extracting or selecting the features before applying a clustering algorithm. It is not obvious to evaluate the importance of the features since the most popular methods to do it are usually made for a supervised learning technique process. A clustering algorithm is an unsupervised method. It means that there is no known output label to match the input data. This paper proposes a new method to choose the best dimensionality reduction method (selection or extraction) according to the data scientist's parameters, aiming to apply a clustering process at the end. It uses Feature Ranking Process Based on Silhouette Decomposition (FRSD) algorithm, a Principal Component Analysis (PCA) algorithm, and a K-Means algorithm along with its metric, the Silhouette Index (SI). This paper presents 5 use cases based on a smart city dataset. This research also aims to discuss the impacts, the advantages, and the disadvantages of each choice that can be made in this unsupervised learning process.

📄 PDF Abstract BibTeX arXiv:2111.10492

Code (0)

등록된 구현이 없습니다.

Tasks

ClusteringDimensionality Reductionfeature selection

Similar Papers 제목 키워드 기반

Randomized Dimensionality Reduction for k-means Clustering

2011-10-13 · Christos Boutsidis, Anastasios Zouzias, Michael W. Mahoney, Petros Drineas

We study the topic of dimensionality reduction for $k$-means clustering. Dimensionality reduction encompasses the union of two approaches: \emph{feature selection} and \emph{feature extraction}. A feature selection based…

ClusteringDimensionality Reductionfeature selection

Feature Selection and Feature Extraction in Pattern Analysis: A Literature Review

2019-05-07 · Benyamin Ghojogh, Maria N. Samad, Sayema Asif Mashhadi, Tania Kapoor 외

Pattern analysis often requires a pre-processing stage for extracting or selecting features in order to help the classification, prediction, or clustering stage discriminate or represent the data in a better way. The rea…

ClusteringDimensionality ReductionFeature EngineeringFeature Importance+2

Unsupervised clustering and classification of upper limb EMG signals during functional movements: a data-driven

2026-05-20 · L. F. Salazar Álvarez, D. Escobar-Saltarén, M. B. Salazar Sánchez, S. C. Henao-Aguirre arxiv

This study presents a comprehensive approach for the clustering and classification of upper-limb surface electromyography (sEMG) signals during functional reach and grasp movements. The methodology was applied to the NIN…

Statistical Parameter Selection for Clustering Persistence Diagrams

2019-10-17 · Max Kontak, Jules Vidal, Julien Tierny

In urgent decision making applications, ensemble simulations are an important way to determine different outcome scenarios based on currently available data. In this paper, we will analyze the output of ensemble simulati…

ClusteringDecision Making

Dimensionality Reduction for $k$-means Clustering

2020-07-26 · Neophytos Charalambides

We present a study on how to effectively reduce the dimensions of the $k$-means clustering problem, so that provably accurate approximations are obtained. Four algorithms are presented, two \textit{feature selection} and…

ClusteringDimensionality Reductionfeature selection