Covariance-based Dissimilarity Measures Applied to Clustering Wide-sense Stationary Ergodic Processes
We introduce a new unsupervised learning problem: clustering wide-sense stationary ergodic stochastic processes. A covariance-based dissimilarity measure together with asymptotically consistent algorithms is designed for clustering offline and online datasets, respectively. We also suggest a formal criterion on the efficiency of dissimilarity measures, and discuss of some approach to improve the efficiency of our clustering algorithms, when they are applied to cluster particular type of processes, such as self-similar processes with wide-sense stationary ergodic increments. Clustering synthetic data and real-world data are provided as examples of applications.
Code (1)
Tasks
ClusteringSimilar Papers 제목 키워드 기반
On Data-Independent Properties for Density-Based Dissimilarity Measures in Hybrid Clustering
Hybrid clustering combines partitional and hierarchical clustering for computational effectiveness and versatility in cluster shape. In such clustering, a dissimilarity measure plays a crucial role in the hierarchical me…
ClusteringImpact of Event Encoding and Dissimilarity Measures on Traffic Crash Characterization Based on Sequence of Events
Crash sequence analysis has been shown in prior studies to be useful for characterizing crashes and identifying safety countermeasures. Sequence analysis is highly domain-specific, but its various techniques have not bee…
ClusteringHierarchical Clustering for Smart Meter Electricity Loads based on Quantile Autocovariances
In order to improve the efficiency and sustainability of electricity systems, most countries worldwide are deploying advanced metering infrastructures, and in particular household smart meters, in the residential sector.…
ClusteringTime SeriesTime Series AnalysisDistance for Functional Data Clustering Based on Smoothing Parameter Commutation
We propose a novel method to determine the dissimilarity between subjects for functional data clustering. Spline smoothing or interpolation is common to deal with data of such type. Instead of estimating the best-represe…
ClusteringMissing ValuesNumerical IntegrationOutlier DetectionMixed-type Distance Shrinkage and Selection for Clustering via Kernel Metric Learning
Distance-based clustering and classification are widely used in various fields to group mixed numeric and categorical data. In many algorithms, a predefined distance measurement is used to cluster data points based on th…
AttributeClusteringMetric Learning