paper-with-me

홈 › Papers

Predictive Overlapping Co-Clustering

2014-03-08 · Chandrima Sarkar, Jaideep Srivastava

In the past few years co-clustering has emerged as an important data mining tool for two way data analysis. Co-clustering is more advantageous over traditional one dimensional clustering in many ways such as, ability to find highly correlated sub-groups of rows and columns. However, one of the overlooked benefits of co-clustering is that, it can be used to extract meaningful knowledge for various other knowledge extraction purposes. For example, building predictive models with high dimensional data and heterogeneous population is a non-trivial task. Co-clusters extracted from such data, which shows similar pattern in both the dimension, can be used for a more accurate predictive model building. Several applications such as finding patient-disease cohorts in health care analysis, finding user-genre groups in recommendation systems and community detection problems can benefit from co-clustering technique that utilizes the predictive power of the data to generate co-clusters for improved data analysis. In this paper, we present the novel idea of Predictive Overlapping Co-Clustering (POCC) as an optimization problem for a more effective and improved predictive analysis. Our algorithm generates optimal co-clusters by maximizing predictive power of the co-clusters subject to the constraints on the number of row and column clusters. In this paper precision, recall and f-measure have been used as evaluation measures of the resulting co-clusters. Results of our algorithm has been compared with two other well-known techniques - K-means and Spectral co-clustering, over four real data set namely, Leukemia, Internet-Ads, Ovarian cancer and MovieLens data set. The results demonstrate the effectiveness and utility of our algorithm POCC in practice.

📄 PDF Abstract BibTeX arXiv:1403.1942

Code (0)

등록된 구현이 없습니다.

Tasks

ClusteringCommunity DetectionRecommendation Systems

Similar Papers 제목 키워드 기반

Overlapping oriented imbalanced ensemble learning method based on projective clustering and stagewise hybrid sampling

2022-11-30 · Fan Li, Bo wang, Pin Wang, Yongming Li

The challenge of imbalanced learning lies not only in class imbalance problem, but also in the class overlapping problem which is complex. However, most of the existing algorithms mainly focus on the former. The limitati…

ClusteringEnsemble LearningTransfer Learning

Element-centric clustering comparison unifies overlaps and hierarchy

2017-06-19 · Alexander J. Gates, Ian B. Wood, William P. Hetrick, Yong-Yeol Ahn

Clustering is one of the most universal approaches for understanding complex data. A pivotal aspect of clustering analysis is quantitatively comparing clusterings; clustering comparison is the basis for many tasks such a…

ClusteringDisentanglementPhilosophy

Functorial Hierarchical Clustering with Overlaps

2016-09-08 · Jared Culbertson, Dan P. Guralnik, Peter F. Stiller

This work draws inspiration from three important sources of research on dissimilarity-based clustering and intertwines those three threads into a consistent principled functorial theory of clustering. Those three are the…

Clustering

Consistency constraints for overlapping data clustering

2016-08-15 · Jared Culbertson, Dan P. Guralnik, Jakob Hansen, Peter F. Stiller

We examine overlapping clustering schemes with functorial constraints, in the spirit of Carlsson--Memoli. This avoids issues arising from the chaining required by partition-based methods. Our principal result shows that …

Clustering

Functorial Clustering via Simplicial Complexes

2020-10-10 · NeurIPS Workshop TDA_and_Beyond 2020 12 · Dan Shiebler

We adapt previous research on topological unsupervised learning to characterize hierarchical overlapping clustering algorithms as functors that factor through a category of simplicial complexes. We first develop a pair o…

Clustering