paper-with-me

홈 › Papers

Ranking and significance of variable-length similarity-based time series motifs

2015-03-06 · Joan Serrà, Isabel Serra, Álvaro Corral, Josep Lluis Arcos

The detection of very similar patterns in a time series, commonly called motifs, has received continuous and increasing attention from diverse scientific communities. In particular, recent approaches for discovering similar motifs of different lengths have been proposed. In this work, we show that such variable-length similarity-based motifs cannot be directly compared, and hence ranked, by their normalized dissimilarities. Specifically, we find that length-normalized motif dissimilarities still have intrinsic dependencies on the motif length, and that lowest dissimilarities are particularly affected by this dependency. Moreover, we find that such dependencies are generally non-linear and change with the considered data set and dissimilarity measure. Based on these findings, we propose a solution to rank those motifs and measure their significance. This solution relies on a compact but accurate model of the dissimilarity space, using a beta distribution with three parameters that depend on the motif length in a non-linear way. We believe the incomparability of variable-length dissimilarities could go beyond the field of time series, and that similar modeling strategies as the one used here could be of help in a more broad context.

📄 PDF Abstract BibTeX arXiv:1503.01883

Code (0)

등록된 구현이 없습니다.

Tasks

Time SeriesTime Series Analysis

Similar Papers 제목 키워드 기반

Self-Supervised Document Similarity Ranking via Contextualized Language Models and Hierarchical Inference

2021-06-02 · Findings (ACL) 2021 8 · Dvir Ginzburg, Itzik Malkiel, Oren Barkan, Avi Caciularu 외

We present a novel model for the problem of ranking a collection of documents according to their semantic similarity to a source (query) document. While the problem of document-to-document similarity ranking has been stu…

Semantic SimilaritySemantic Textual Similarity

RAxSS: Retrieval-Augmented Sparse Sampling for Explainable Variable-Length Medical Time Series Classification

2025-10-03 · Aydin Javadov, Samir Garibov, Tobias Hoesli, Qiyang Sun 외 arxiv

Medical time series analysis is challenging due to data sparsity, noise, and highly variable recording lengths. Prior work has shown that stochastic sparse sampling effectively handles variable-length signals, while retr…

Time Series ClassificationTime Series Analysis

SummerTime: Variable-length Time SeriesSummarization with Applications to PhysicalActivity Analysis

2020-02-20 · Kevin M. Amaral, Zihan Li, Wei Ding, Scott Crouter 외

\textit{SummerTime} seeks to summarize globally time series signals and provides a fixed-length, robust summarization of the variable-length time series. Many classical machine learning methods for classification and reg…

General ClassificationregressionTime SeriesTime Series Analysis

EFloat: Entropy-coded Floating Point Format for Compressing Vector Embedding Models

2021-02-04 · NeurIPS 2021 12 · Rajesh Bordawekar, Bulent Abali, Ming-Hung Chen

In a large class of deep learning models, including vector embedding models such as word and database embeddings, we observe that floating point exponent values cluster around a few unique values, permitting entropy base…

Data Compression

Sequential Learning-based IaaS Composition

2021-02-24 · Sajib Mistry, Sheik Mohammad Mostakim Fattah, Athman Bouguettaya

We propose a novel IaaS composition framework that selects an optimal set of consumer requests according to the provider's qualitative preferences on long-term service provisions. Decision variables are included in the t…

ClusteringQ-Learning