paper-with-me

홈 › Papers

A Comparison of Polynomial-Based Tree Clustering Methods

2026-01-13 · Pengyu Liu, Mariel Vázquez, Nataša Jonoska arxiv

Tree structures appear in many fields of the life sciences, including phylogenetics, developmental biology and nucleic acid structures. Trees can be used to represent RNA secondary structures, which directly relate to the function of non-coding RNAs. Recent developments in sequencing technology and artificial intelligence have yielded numerous biological data that can be represented with tree structures. This requires novel methods for tree structure data analytics. Tree polynomials provide a computationally efficient, interpretable and comprehensive way to encode tree structures as matrices, which are compatible with most data analytics tools. Machine learning methods based on the Canberra distance between tree polynomials have been introduced to analyze phylogenies and nucleic acid structures. In this paper, we compare the performance of different distances in tree clustering methods based on a tree distinguishing polynomial. We also implement two basic autoencoder models for clustering trees using the polynomial. We find that the distance based methods with entry-level normalized distances have the highest clustering accuracy among the compared methods.

📄 PDF Abstract BibTeX arXiv:2601.14285

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

An objective function for order preserving hierarchical clustering

2021-09-09 · Daniel Bakkelund

We present a theory and an objective function for similarity-based hierarchical clustering of probabilistic partial orders and directed acyclic graphs (DAGs). Specifically, given elements $x \le y$ in the partial order, …

ClusteringRelation

Learning-Augmented Hierarchical Clustering

2025-06-05 · Vladimir Braverman, Jon C. Ergun, Chen Wang, Samson Zhou

Hierarchical clustering (HC) is an important data analysis technique in which the goal is to recursively partition a dataset into a tree-like structure while grouping together similar data points at each level of granula…

ClusteringTriplet

Characterizing Admissible Objective Functions for Hierarchical Clustering

2026-04-26 · Ryuki Tsukuba, Kazutoshi Ando arxiv

Hierarchical clustering is a fundamental task in data analysis, but classical methods have long lacked a principled objective function. Dasgupta [STOC 2016] took an important step toward addressing this gap by proposing …

The computational complexity of some explainable clustering problems

2022-08-20 · Eduardo Sany Laber

We study the computational complexity of some explainable clustering problems in the framework proposed by [Dasgupta et al., ICML 2020], where explainability is achieved via axis-aligned decision trees. We consider the $…

Clustering

Model-based clustering with Hidden Markov Model regression for time series with regime changes

2013-12-25 · Faicel Chamroukhi, Allou Samé, Patrice Aknin, Gérard Govaert

This paper introduces a novel model-based clustering approach for clustering time series which present changes in regime. It consists of a mixture of polynomial regressions governed by hidden Markov chains. The underlyin…

Clusteringmodelparameter estimationregression+3