paper-with-me

Papers

Sparse group factor analysis for biclustering of multiple data sources

2015-12-29 · Kerstin Bunte, Eemeli Leppäaho, Inka Saarinen, Samuel Kaski

Motivation: Modelling methods that find structure in data are necessary with the current large volumes of genomic data, and there have been various efforts to find subsets of genes exhibiting consistent patterns over subsets of treatments. These biclustering techniques have focused on one data source, often gene expression data. We present a Bayesian approach for joint biclustering of multiple data sources, extending a recent method Group Factor Analysis (GFA) to have a biclustering interpretation with additional sparsity assumptions. The resulting method enables data-driven detection of linear structure present in parts of the data sources. Results: Our simulation studies show that the proposed method reliably infers bi-clusters from heterogeneous data sources. We tested the method on data from the NCI-DREAM drug sensitivity prediction challenge, resulting in an excellent prediction accuracy. Moreover, the predictions are based on several biclusters which provide insight into the data sources, in this case on gene expression, DNA methylation, protein abundance, exome sequence, functional connectivity fingerprints and drug sensitivity.

📄 PDF Abstract BibTeX arXiv:1512.08808

Code (0)

등록된 구현이 없습니다.

Tasks

Functional ConnectivitySensitivity

Similar Papers 제목 키워드 기반

Biclustering Methods via Sparse Penalty

2023-08-28 · Jiqiang Wang

In this paper, we first reviewed several biclustering methods that are used to identify the most significant clusters in gene expression data. Here we mainly focused on the SSVD(sparse SVD) method and tried a new sparse …

Minimax Localization of Structural Information in Large Noisy Matrices

2011-12-01 · NeurIPS 2011 12 · Mladen Kolar, Sivaraman Balakrishnan, Alessandro Rinaldo, Aarti Singh

We consider the problem of identifying a sparse set of relevant columns and rows in a large data matrix with highly corrupted entries. This problem of identifying groups from a collection of bipartite variables such as p…

ClusteringTwo-sample testing

Sparse Convex Biclustering

2026-01-05 · Jiakun Jiang, Dewei Xiang, Chenliang Gu, Wei Liu 외 arxiv

Biclustering is an essential unsupervised machine learning technique for simultaneously clustering rows and columns of a data matrix, with widespread applications in genomics, transcriptomics, and other high-dimensional …

Towards a Unified Taxonomy of Biclustering Methods

2017-02-17 · Dmitry I. Ignatov, Bruce W. Watson

Being an unsupervised machine learning and data mining technique, biclustering and its multimodal extensions are becoming popular tools for analysing object-attribute data in different domains. Apart from conventional cl…

AttributeClusteringSurvey

SOFAR: large-scale association network learning

2017-04-26 · Yoshimasa Uematsu, Yingying Fan, Kun Chen, Jinchi Lv 외

Many modern big data applications feature large scale in both numbers of responses and predictors. Better statistical efficiency and scientific insights can be enabled by understanding the large-scale response-predictor …