paper-with-me

홈 › Papers

Towards Generalisable Time Series Understanding Across Domains

2024-10-09 · Özgün Turgut, Philip Müller, Martin J. Menten, Daniel Rueckert

Recent breakthroughs in natural language processing and computer vision, driven by efficient pre-training on large datasets, have enabled foundation models to excel on a wide range of tasks. However, this potential has not yet been fully realised in time series analysis, as existing methods fail to address the heterogeneity in large time series corpora. Prevalent in domains ranging from medicine to finance, time series vary substantially in characteristics such as variate count, inter-variate relationships, temporal patterns, and sampling frequency. To address this, we introduce a novel pre-training paradigm specifically designed to handle time series heterogeneity. We propose a tokeniser with learnable domain signatures, a dual masking strategy, and a normalised cross-correlation loss, enabling our open model for general time series analysis (OTiS) to efficiently learn from large time series corpora. Extensive benchmarking on diverse tasks, such as classification, regression, and forecasting, demonstrates that OTiS outperforms state-of-the-art baselines. Our code and pre-trained weights are available at https://github.com/oetu/otis.

📄 PDF Abstract BibTeX arXiv:2410.07299

Code (1)

oetu/otis 공식 구현 pytorch

Tasks

BenchmarkingTime SeriesTime Series Analysis

Similar Papers 제목 키워드 기반

NLU++: A Multi-Label, Slot-Rich, Generalisable Dataset for Natural Language Understanding in Task-Oriented Dialogue

2022-04-27 · Findings (NAACL) 2022 7 · Iñigo Casanueva, Ivan Vulić, Georgios P. Spithourakis, Paweł Budzianowski

We present NLU++, a novel dataset for natural language understanding (NLU) in task-oriented dialogue (ToD) systems, with the aim to provide a much more challenging evaluation environment for dialogue NLU models, up to da…

Natural Language Understanding

How does downsampling affect needle electromyography signals? A generalisable workflow for understanding downsampling effects on high-frequency time series

2026-01-15 · Mathieu Cherpitel, Janne Luijten, Thomas Bäck, Camiel Verhamme 외 arxiv

Automated analysis of needle electromyography (nEMG) signals is emerging as a tool to support the detection of neuromuscular diseases (NMDs), yet the signals' high and heterogeneous sampling rates pose substantial comput…

Towards Time Series Reasoning with LLMs

2024-09-17 · Winnie Chow, Lauren Gardiner, Haraldur T. Hallgrímsson, Maxwell A. Xu 외

Multi-modal large language models (MLLMs) have enabled numerous advances in understanding and reasoning in domains like vision, but we have not yet seen this broad success for time-series. Although prior works on time-se…

Time SeriesTime Series Forecasting

DoDo Learning: DOmain-DemOgraphic Transfer in Language Models for Detecting Abuse Targeted at Public Figures

2023-07-31 · Angus R. Williams, Hannah Rose Kirk, Liam Burke, Yi-Ling Chung 외

Public figures receive a disproportionate amount of abuse on social media, impacting their active participation in public life. Automated systems can identify abuse at scale but labelling training data is expensive, comp…

text-classificationText Classification

GETA: Generalized Encrypted Traffic Analysis

2026-05-29 · Ransika Gunasekara, Rahat Masood, Salil Kanhere arxiv

Traditional traffic analysis is being fundamentally challenged by the rapid adoption of encryption, tunnelling, and privacy-preserving protocols, which increasingly obscure packet payloads and limit the usefulness of Dee…