paper-with-me

Papers

Ranked differences Pearson correlation dissimilarity with an application to electricity users time series clustering

2025-05-04 · Chutiphan Charoensuk, Nathakhun Wiroonsri

Time series clustering is an unsupervised learning method for classifying time series data into groups with similar behavior. It is used in applications such as healthcare, finance, economics, energy, and climate science. Several time series clustering methods have been introduced and used for over four decades. Most of them focus on measuring either Euclidean distances or association dissimilarities between time series. In this work, we propose a new dissimilarity measure called ranked Pearson correlation dissimilarity (RDPC), which combines a weighted average of a specified fraction of the largest element-wise differences with the well-known Pearson correlation dissimilarity. It is incorporated into hierarchical clustering. The performance is evaluated and compared with existing clustering algorithms. The results show that the RDPC algorithm outperforms others in complicated cases involving different seasonal patterns, trends, and peaks. Finally, we demonstrate our method by clustering a random sample of customers from a Thai electricity consumption time series dataset into seven groups with unique characteristics.

📄 PDF Abstract BibTeX arXiv:2505.02173

Code (0)

등록된 구현이 없습니다.

Tasks

ClusteringTime SeriesTime Series Clustering

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Don't Sweat the Small Stuff: Segment-Level Meta-Evaluation Based on Pairwise Difference Correlation

2025-09-29 · Colten DiIanni, Daniel Deutsch arxiv

This paper introduces Pairwise Difference Pearson (PDP), a novel segment-level meta-evaluation metric for Machine Translation (MT) that address limitations in previous Pearson's $ρ$-based and and Kendall's $τ$-based meta…

Machine Translation

Gower's similarity coefficients with automatic weight selection

2024-01-30 · Marcello D'Orazio

Nearest-neighbor methods have become popular in statistics and play a key role in statistical learning. Important decisions in nearest-neighbor methods concern the variables to use (when many potential candidates exist) …

ImputationMissing Values

Matrix dissimilarities based on differences in moments and sparsity

2024-06-04 · Li Tuobang

Generating a dissimilarity matrix is typically the first step in big data analysis. Although numerous methods exist, such as Euclidean distance, Minkowski distance, Manhattan distance, Bray Curtis dissimilarity, Jaccard …

Predicting Psychological Health from Childhood Essays with Convolutional Neural Networks for the CLPsych 2018 Shared Task (Team UKNLP)

2018-06-01 · WS 2018 6 · Anthony Rios, Tung Tran, Ramakanth Kavuluru

This paper describes the systems we developed for tasks A and B of the 2018 CLPsych shared task. The first task (task A) focuses on predicting behavioral health scores at age 11 using childhood essays. The second task (t…

regression

Cross-Correlation Based Discriminant Criterion for Channel Selection in Motor Imagery BCI Systems

2020-12-03 · Jianli Yu, Zhuliang Yu

Objective. Many electroencephalogram (EEG)-based brain-computer interface (BCI) systems use a large amount of channels for higher performance, which is time-consuming to set up and inconvenient for practical applications…

Brain Computer Interfacechannel selectionEEGElectroencephalogram (EEG)+1