Time series clustering based on the characterisation of segment typologies
Time series clustering is the process of grouping time series with respect to their similarity or characteristics. Previous approaches usually combine a specific distance measure for time series and a standard clustering method. However, these approaches do not take the similarity of the different subsequences of each time series into account, which can be used to better compare the time series objects of the dataset. In this paper, we propose a novel technique of time series clustering based on two clustering stages. In a first step, a least squares polynomial segmentation procedure is applied to each time series, which is based on a growing window technique that returns different-length segments. Then, all the segments are projected into same dimensional space, based on the coefficients of the model that approximates the segment and a set of statistical features. After mapping, a first hierarchical clustering phase is applied to all mapped segments, returning groups of segments for each time series. These clusters are used to represent all time series in the same dimensional space, after defining another specific mapping process. In a second and final clustering stage, all the time series objects are grouped. We consider internal clustering quality to automatically adjust the main parameter of the algorithm, which is an error threshold for the segmenta- tion. The results obtained on 84 datasets from the UCR Time Series Classification Archive have been compared against two state-of-the-art methods, showing that the performance of this methodology is very promising.
Code (0)
등록된 구현이 없습니다.
Tasks
ClusteringTime SeriesTime Series AnalysisTime Series ClassificationTime Series ClusteringSimilar Papers 제목 키워드 기반
AMP: a new time-frequency feature extraction method for intermittent time-series data
The characterisation of time-series data via their most salient features is extremely important in a range of machine learning task, not least of all with regards to classification and clustering. While there exist many …
ClusteringTime SeriesTime Series AnalysisClustering of Urban Traffic Patterns by K-Means and Dynamic Time Warping: Case Study
Clustering of urban traffic patterns is an essential task in many different areas of traffic management and planning. In this paper, two significant applications in the clustering of urban traffic patterns are described.…
ClusteringDynamic Time WarpingManagementTime Series+1Optimal Transport Based Change Point Detection and Time Series Segment Clustering
Two common problems in time series analysis are the decomposition of the data stream into disjoint segments that are each in some sense "homogeneous" - a problem known as Change Point Detection (CPD) - and the grouping o…
Change Point DetectionClusteringTime SeriesTime Series AnalysisdhSegment: A generic deep-learning approach for document segmentation
In recent years there have been multiple successful attempts tackling document processing problems separately by designing task specific hand-tuned strategies. We argue that the diversity of historical document processin…
Deep LearningDiversityDocument Layout AnalysisClustering Time-Series by a Novel Slope-Based Similarity Measure Considering Particle Swarm Optimization
Recently there has been an increase in the studies on time-series data mining specifically time-series clustering due to the vast existence of time-series in various domains. The large volume of data in the form of time-…
ClusteringDynamic Time WarpingTime SeriesTime Series Analysis+1