paper-with-me

Papers

Fast Data Driven Estimation of Cluster Number in Multiplex Images using Embedded Density Outliers

2022-07-21 · Spencer A. Thomas

The usage of chemical imaging technologies is becoming a routine accompaniment to traditional methods in pathology. Significant technological advances have developed these next generation techniques to provide rich, spatially resolved, multidimensional chemical images. The rise of digital pathology has significantly enhanced the synergy of these imaging modalities with optical microscopy and immunohistochemistry, enhancing our understanding of the biological mechanisms and progression of diseases. Techniques such as imaging mass cytometry provide labelled multidimensional (multiplex) images of specific components used in conjunction with digital pathology techniques. These powerful techniques generate a wealth of high dimensional data that create significant challenges in data analysis. Unsupervised methods such as clustering are an attractive way to analyse these data, however, they require the selection of parameters such as the number of clusters. Here we propose a methodology to estimate the number of clusters in an automatic data-driven manner using a deep sparse autoencoder to embed the data into a lower dimensional space. We compute the density of regions in the embedded space, the majority of which are empty, enabling the high density regions to be detected as outliers and provide an estimate for the number of clusters. This framework provides a fully unsupervised and data-driven method to analyse multidimensional data. In this work we demonstrate our method using 45 multiplex imaging mass cytometry datasets. Moreover, our model is trained using only one of the datasets and the learned embedding is applied to the remaining 44 images providing an efficient process for data analysis. Finally, we demonstrate the high computational efficiency of our method which is two orders of magnitude faster than estimating via computing the sum squared distances as a function of cluster number.

📄 PDF Abstract BibTeX arXiv:2207.10469

Code (0)

등록된 구현이 없습니다.

Tasks

Computational Efficiency

Methods 이 논문이 사용한 방법론

Sparse Autoencoder A Sparse Autoencoder is a type of autoencoder that employs sparsity to achieve an information bottleneck. Specifically the loss function is constructed so that activations are…

Similar Papers 제목 키워드 기반

Finding Geometric Models by Clustering in the Consensus Space

2021-03-25 · CVPR 2023 1 · Daniel Barath, Denys Rozumny, Ivan Eichhardt, Levente Hajder 외

We propose a new algorithm for finding an unknown number of geometric models, e.g., homographies. The problem is formalized as finding dominant model instances progressively without forming crisp point-to-model assignmen…

ClusteringMotion EstimationPose Estimation

Electrode Clustering and Bandpass Analysis of EEG Data for Gaze Estimation

2023-02-19 · Ard Kastrati, Martyna Beata Plomecka, Joël Küchler, Nicolas Langer 외

In this study, we validate the findings of previously published papers, showing the feasibility of an Electroencephalography (EEG) based gaze estimation. Moreover, we extend previous research by demonstrating that with o…

ClusteringEEGElectroencephalogram (EEG)Gaze Estimation

Efficient, Effective and Well Justified Estimation of Active Nodes within a Cluster

2020-01-26 · Md Mahmudul Hasan, Shuangqing Wei, Ramachandran Vaidyanathan

Reliable and efficient estimation of the size of a dynamically changing cluster in an IoT network is critical in its nominal operation. Most previous estimation schemes worked with relatively smaller frame size and large…

Sparse Bayesian Unsupervised Learning

2014-01-30 · Stephane Gaiffas, Bertrand Michel

This paper is about variable selection, clustering and estimation in an unsupervised high-dimensional setting. Our approach is based on fitting constrained Gaussian mixture models, where we learn the number of clusters $…

ClusteringVariable Selection

Bayesian temporal biclustering with applications to multi-subject neuroscience studies

2024-06-24 · Federica Zoe Ricci, Erik B. Sudderth, Jaylen Lee, Megan A. K. Peters 외

We consider the problem of analyzing multivariate time series collected on multiple subjects, with the goal of identifying groups of subjects exhibiting similar trends in their recorded measurements over time as well as …