paper-with-me

홈 › Papers

Subject clustering by IF-PCA and several recent methods

2023-06-08 · Dieyi Chen, Jiashun Jin, Zheng Tracy Ke

Subject clustering (i.e., the use of measured features to cluster subjects, such as patients or cells, into multiple groups) is a problem of great interest. In recent years, many approaches were proposed, among which unsupervised deep learning (UDL) has received a great deal of attention. Two interesting questions are (a) how to combine the strengths of UDL and other approaches, and (b) how these approaches compare to one other. We combine Variational Auto-Encoder (VAE), a popular UDL approach, with the recent idea of Influential Feature PCA (IF-PCA), and propose IF-VAE as a new method for subject clustering. We study IF-VAE and compare it with several other methods (including IF-PCA, VAE, Seurat, and SC3) on $10$ gene microarray data sets and $8$ single-cell RNA-seq data sets. We find that IF-VAE significantly improves over VAE, but still underperforms IF-PCA. We also find that IF-PCA is quite competitive, which slightly outperforms Seurat and SC3 over the $8$ single-cell data sets. IF-PCA is conceptually simple and permits delicate analysis. We demonstrate that IF-PCA is capable of achieving the phase transition in a Rare/Weak model. Comparatively, Seurat and SC3 are more complex and theoretically difficult to analyze (for these reasons, their optimality remains unclear).

📄 PDF Abstract BibTeX arXiv:2306.05363

Code (1)

zhengtracyke/ifpca 공식 구현

Tasks

Clustering

Methods 이 논문이 사용한 방법론

PCA Principle Components Analysis (PCA) is an unsupervised method primary used for dimensionality reduction within machine learning. PCA is calculated via a singular value…

Similar Papers 제목 키워드 기반

A Pragmatic Method for Comparing Clusterings with Overlaps and Outliers

2026-02-16 · Ryan DeWolfe, Paweł Prałat, François Théberge arxiv

Clustering algorithms are an essential part of the unsupervised data science ecosystem, and extrinsic evaluation of clustering algorithms requires a method for comparing the detected clustering to a ground truth clusteri…

Multiple-View Spectral Clustering for Group-wise Functional Community Detection

2016-11-21 · Nathan D. Cahill, Harmeet Singh, Chao Zhang, Daryl A. Corcoran 외

Functional connectivity analysis yields powerful insights into our understanding of the human brain. Group-wise functional community detection aims to partition the brain into clusters, or communities, in which functiona…

ClusteringCommunity DetectionFunctional Connectivity

Distance for Functional Data Clustering Based on Smoothing Parameter Commutation

2016-04-10 · ShengLi Tzeng, Christian Hennig, Yu-Fen Li, Chien-Ju Lin

We propose a novel method to determine the dissimilarity between subjects for functional data clustering. Spline smoothing or interpolation is common to deal with data of such type. Instead of estimating the best-represe…

ClusteringMissing ValuesNumerical IntegrationOutlier Detection

Explicit agreement extremes for a $2\times2$ table with given marginals

2020-01-21 · José E. Chacón

The problem of maximizing (or minimizing) the agreement between clusterings, subject to given marginals, can be formally posed under a common framework for several agreement measures. Until now, it was possible to find i…

Subject Enveloped Deep Sample Fuzzy Ensemble Learning Algorithm of Parkinson's Speech Data

2021-11-17 · Yiwen Wang, Fan Li, Xiaoheng Zhang, Pin Wang 외

Parkinson disease (PD)'s speech recognition is an effective way for its diagnosis, which has become a hot and difficult research area in recent years. As we know, there are large corpuses (segments) within one subject. H…

DiagnosticEnsemble Learningspeech-recognitionSpeech Recognition