paper-with-me

Papers

On Data-Independent Properties for Density-Based Dissimilarity Measures in Hybrid Clustering

2016-09-21 · Kajsa Møllersen, Subhra S. Dhar, Fred Godtliebsen

Hybrid clustering combines partitional and hierarchical clustering for computational effectiveness and versatility in cluster shape. In such clustering, a dissimilarity measure plays a crucial role in the hierarchical merging. The dissimilarity measure has great impact on the final clustering, and data-independent properties are needed to choose the right dissimilarity measure for the problem at hand. Properties for distance-based dissimilarity measures have been studied for decades, but properties for density-based dissimilarity measures have so far received little attention. Here, we propose six data-independent properties to evaluate density-based dissimilarity measures associated with hybrid clustering, regarding equality, orthogonality, symmetry, outlier and noise observations, and light-tailed models for heavy-tailed clusters. The significance of the properties is investigated, and we study some well-known dissimilarity measures based on Shannon entropy, misclassification rate, Bhattacharyya distance and Kullback-Leibler divergence with respect to the proposed properties. As none of them satisfy all the proposed properties, we introduce a new dissimilarity measure based on the Kullback-Leibler information and show that it satisfies all proposed properties. The effect of the proposed properties is also illustrated on several real and simulated data sets.

📄 PDF Abstract BibTeX arXiv:1609.06533

Code (0)

등록된 구현이 없습니다.

Tasks

Clustering

Similar Papers 제목 키워드 기반

Understanding the Inner Workings of Language Models Through Representation Dissimilarity

2023-10-23 · Davis Brown, Charles Godfrey, Nicholas Konz, Jonathan Tu 외

As language models are applied to an increasing number of real-world applications, understanding their inner workings has become an important issue in model trust, interpretability, and transparency. In this work we show…

Language ModelingLanguage Modelling

Metric Space Magnitude for Evaluating the Diversity of Latent Representations

2023-11-27 · Katharina Limbeck, Rayna Andreeva, Rik Sarkar, Bastian Rieck

The magnitude of a metric space is a novel invariant that provides a measure of the 'effective size' of a space across multiple scales, while also capturing numerous geometrical properties, such as curvature, density, or…

Dimensionality ReductionDiversityRepresentation Learning

Efficient Bregman Range Search

2009-12-01 · NeurIPS 2009 12 · Lawrence Cayton

We develop an algorithm for efficient range search when the notion of dissimilarity is given by a Bregman divergence. The range search task is to return all points in a potentially large database that are within some sp…

Density EstimationInformation RetrievalOutlier DetectionRetrieval

A guide through a family of phylogenetic dissimilarity measures among sites

2015-06-21

Ecological studies have now gone beyond measures of species turnover towards measures of phylogenetic and functional dissimilarity with a main objective: disentangling the processes that drive species distributions from …

Diversity

Towards an Axiomatic Approach to Hierarchical Clustering of Measures

2015-08-15 · Philipp Thomann, Ingo Steinwart, Nico Schmid

We propose some axioms for hierarchical clustering of probability measures and investigate their ramifications. The basic idea is to let the user stipulate the clusters for some elementary measures. This is done without …

Clustering