paper-with-me

Papers

Consistent spectral clustering in sparse tensor block models

2025-01-23 · Ian Välimaa, Lasse Leskelä

High-order clustering aims to classify objects in multiway datasets that are prevalent in various fields such as bioinformatics, social network analysis, and recommendation systems. These tasks often involve data that is sparse and high-dimensional, presenting significant statistical and computational challenges. This paper introduces a tensor block model specifically designed for sparse integer-valued data tensors. We propose a simple spectral clustering algorithm augmented with a trimming step to mitigate noise fluctuations, and identify a density threshold that ensures the algorithm's consistency. Our approach models sparsity using a sub-Poisson noise concentration framework, accommodating heavier than sub-Gaussian tails. Remarkably, this natural class of tensor block models is closed under aggregation across arbitrary modes. Consequently, we obtain a comprehensive framework for evaluating the tradeoff between signal loss and noise reduction during data aggregation. The analysis is based on a novel concentration bound for sparse random Gram matrices. The theoretical findings are illustrated through simulation experiments.

📄 PDF Abstract BibTeX arXiv:2501.13820

Code (0)

등록된 구현이 없습니다.

Tasks

ClusteringRecommendation Systems

Methods 이 논문이 사용한 방법론

Spectral Clustering Spectral clustering has attracted increasing attention due to the promising ability in dealing with nonlinearly separable datasets [15], [16]. In spectral clustering, the…

Similar Papers 제목 키워드 기반

Sparse Subspace Clustering in Diverse Multiplex Network Model

2022-06-15 · Majid Noroozi, Marianna Pensky

The paper considers the DIverse MultiPLEx (DIMPLE) network model, introduced in Pensky and Wang (2021), where all layers of the network have the same collection of nodes and are equipped with the Stochastic Block Models.…

ClusteringmodelStochastic Block Model

Exact Clustering in Tensor Block Model: Statistical Optimality and Computational Limit

2020-12-18 · Rungang Han, Yuetian Luo, Miaoyan Wang, Anru R. Zhang

High-order clustering aims to identify heterogeneous substructures in multiway datasets that arise commonly in neuroimaging, genomics, social network studies, etc. The non-convex and discontinuous nature of this problem …

Clustering

Sparse and Smooth: improved guarantees for Spectral Clustering in the Dynamic Stochastic Block Model

2020-02-07 · Nicolas Keriven, Samuel Vaiter

In this paper, we analyse classical variants of the Spectral Clustering (SC) algorithm in the Dynamic Stochastic Block Model (DSBM). Existing results show that, in the relatively sparse case where the expected degree gro…

ClusteringStochastic Block Model

Pseudo-likelihood methods for community detection in large sparse networks

2012-07-10 · Arash A. Amini, Aiyou Chen, Peter J. Bickel, Elizaveta Levina

Many algorithms have been proposed for fitting network models with communities, but most of them do not scale well to large networks, and often fail on sparse networks. Here we propose a new fast pseudo-likelihood method…

ClusteringCommunity DetectionStochastic Block Model

Tensor Sparse and Low-Rank based Submodule Clustering Method for Multi-way Data

2016-01-02 · Xinglin Piao, Yongli Hu, Junbin Gao, Yanfeng Sun 외

A new submodule clustering method via sparse and low-rank representation for multi-way data is proposed in this paper. Instead of reshaping multi-way data into vectors, this method maintains their natural orders to prese…

Clustering