paper-with-me

홈 › Papers

Expectation Distance-based Distributional Clustering for Noise-Robustness

2021-10-17 · Rahmat Adesunkanmi, Ratnesh Kumar

This paper presents a clustering technique that reduces the susceptibility to data noise by learning and clustering the data-distribution and then assigning the data to the cluster of its distribution. In the process, it reduces the impact of noise on clustering results. This method involves introducing a new distance among distributions, namely the expectation distance (denoted, ED), that goes beyond the state-of-art distribution distance of optimal mass transport (denoted, $W_2$ for $2$-Wasserstein): The latter essentially depends only on the marginal distributions while the former also employs the information about the joint distributions. Using the ED, the paper extends the classical $K$-means and $K$-medoids clustering to those over data-distributions (rather than raw-data) and introduces $K$-medoids using $W_2$. The paper also presents the closed-form expressions of the $W_2$ and ED distance measures. The implementation results of the proposed ED and the $W_2$ distance measures to cluster real-world weather data as well as stock data are also presented, which involves efficiently extracting and using the underlying data distributions -- Gaussians for weather data versus lognormals for stock data. The results show striking performance improvement over classical clustering of raw-data, with higher accuracy realized for ED. Also, not only does the distribution-based clustering offer higher accuracy, but it also lowers the computation time due to reduced time-complexity.

📄 PDF Abstract BibTeX arXiv:2110.08871

Code (0)

등록된 구현이 없습니다.

Tasks

Clustering

Similar Papers 제목 키워드 기반

Taking a Moment for Distributional Robustness

2024-05-08 · Jabari Hastings, Christopher Jung, Charlotte Peale, Vasilis Syrgkanis

A rich line of recent work has studied distributionally robust learning approaches that seek to learn a hypothesis that performs well, in the worst-case, on many different distributions over a population. We argue that a…

Exploring the Robustness of Distributional Reinforcement Learning against Noisy State Observations

2021-09-29 · Ke Sun, Yi Liu, Yingnan Zhao, Hengshuai Yao 외

In real scenarios, state observations that an agent observes may contain measurement errors or adversarial noises, misleading the agent to take suboptimal actions or even collapse while training. In this paper, we study …

Distributional Reinforcement Learningreinforcement-learningReinforcement Learning (RL)

Exploring the Training Robustness of Distributional Reinforcement Learning against Noisy State Observations

2021-09-17 · Ke Sun, Yingnan Zhao, Shangling Jui, Linglong Kong

In real scenarios, state observations that an agent observes may contain measurement errors or adversarial noises, misleading the agent to take suboptimal actions or even collapse while training. In this paper, we study …

Density EstimationDistributional Reinforcement Learningreinforcement-learningReinforcement Learning (RL)

Distributionally Robust K-Means Clustering

2026-04-13 · Vikrant Malik, Taylan Kargin, Babak Hassibi arxiv

K-means clustering is a workhorse of unsupervised learning, but it is notoriously brittle to outliers, distribution shifts, and limited sample sizes. Viewing k-means as Lloyd--Max quantization of the empirical distributi…

Outlier Detection

Fuzzy clustering of distribution-valued data using adaptive L2 Wasserstein distances

2016-05-02 · Antonio Irpino, Francisco De Carvalho, Rosanna Verde

Distributional (or distribution-valued) data are a new type of data arising from several sources and are considered as realizations of distributional variables. A new set of fuzzy c-means algorithms for data described by…

ClusteringVariable Selection