paper-with-me

Papers

Sparse Convex Biclustering

2026-01-05 · Jiakun Jiang, Dewei Xiang, Chenliang Gu, Wei Liu, Binhuan Wang arxiv

Biclustering is an essential unsupervised machine learning technique for simultaneously clustering rows and columns of a data matrix, with widespread applications in genomics, transcriptomics, and other high-dimensional omics data. Despite its importance, existing biclustering methods struggle to meet the demands of modern large-scale datasets. The challenges stem from the accumulation of noise in high-dimensional features, the limitations of non-convex optimization formulations, and the computational complexity of identifying meaningful biclusters. These issues often result in reduced accuracy and stability as the size of the dataset increases. To overcome these challenges, we propose Sparse Convex Biclustering (SpaCoBi), a novel method that penalizes noise during the biclustering process to improve both accuracy and robustness. By adopting a convex optimization framework and introducing a stability-based tuning criterion, SpaCoBi achieves an optimal balance between cluster fidelity and sparsity. Comprehensive numerical studies, including simulations and an application to mouse olfactory bulb data, demonstrate that SpaCoBi significantly outperforms state-of-the-art methods in accuracy. These results highlight SpaCoBi as a robust and efficient solution for biclustering in high-dimensional and large-scale datasets.

📄 PDF Abstract BibTeX arXiv:2601.01757

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Convex Biclustering

2014-08-05 · Eric C. Chi, Genevera I. Allen, Richard G. Baraniuk

In the biclustering problem, we seek to simultaneously group observations and features. While biclustering has applications in a wide array of domains, ranging from text mining to collaborative filtering, the problem of …

Collaborative Filtering

Robust convex biclustering with a tuning-free method

2022-12-06 · Yifan Chen, Chunyin Lei, Chuanquan Li, Haiqiang Ma 외

Biclustering is widely used in different kinds of fields including gene information analysis, text mining, and recommendation system by effectively discovering the local correlation between samples and features. However,…

Minimax Localization of Structural Information in Large Noisy Matrices

2011-12-01 · NeurIPS 2011 12 · Mladen Kolar, Sivaraman Balakrishnan, Alessandro Rinaldo, Aarti Singh

We consider the problem of identifying a sparse set of relevant columns and rows in a large data matrix with highly corrupted entries. This problem of identifying groups from a collection of bipartite variables such as p…

ClusteringTwo-sample testing

SOFAR: large-scale association network learning

2017-04-26 · Yoshimasa Uematsu, Yingying Fan, Kun Chen, Jinchi Lv 외

Many modern big data applications feature large scale in both numbers of responses and predictors. Better statistical efficiency and scientific insights can be enabled by understanding the large-scale response-predictor …

L0-norm Sparse Graph-regularized SVD for Biclustering

2016-03-19 · Wenwen Min, Juan Liu, Shihua Zhang

Learning the "blocking" structure is a central challenge for high dimensional data (e.g., gene expression data). Recently, a sparse singular value decomposition (SVD) has been used as a biclustering tool to achieve this …

Blocking