paper-with-me

홈 › Papers

GIFT-Eval: A Benchmark For General Time Series Forecasting Model Evaluation

2024-10-14 · Taha Aksu, Gerald Woo, Juncheng Liu, Xu Liu, Chenghao Liu, Silvio Savarese, Caiming Xiong, Doyen Sahoo

Time series foundation models excel in zero-shot forecasting, handling diverse tasks without explicit training. However, the advancement of these models has been hindered by the lack of comprehensive benchmarks. To address this gap, we introduce the General Time Series Forecasting Model Evaluation, GIFT-Eval, a pioneering benchmark aimed at promoting evaluation across diverse datasets. GIFT-Eval encompasses 23 datasets over 144,000 time series and 177 million data points, spanning seven domains, 10 frequencies, multivariate inputs, and prediction lengths ranging from short to long-term forecasts. To facilitate the effective pretraining and evaluation of foundation models, we also provide a non-leaking pretraining dataset containing approximately 230 billion data points. Additionally, we provide a comprehensive analysis of 17 baselines, which includes statistical models, deep learning models, and foundation models. We discuss each model in the context of various benchmark characteristics and offer a qualitative analysis that spans both deep learning and foundation models. We believe the insights from this analysis, along with access to this new standard zero-shot time series forecasting benchmark, will guide future developments in time series foundation models. Code, data, and the leaderboard can be found at https://github.com/SalesforceAIResearch/gift-eval .

📄 PDF Abstract BibTeX arXiv:2410.10393

Code (1)

salesforceairesearch/gift-eval 공식 구현

Tasks

Time SeriesTime Series Forecasting

Similar Papers 제목 키워드 기반

Cisco Time Series Model Technical Report

2025-11-25 · Liang Gou, Archit Khare, Praneet Pabolu, Prachi Patel 외 arxiv

We introduce the Cisco Time Series Model, a univariate zero-shot forecaster. This time series foundation model is the result of a general architectural innovation to a time series model enabling it to accept multiresolut…

Toto 2.0: Time Series Forecasting Enters the Scaling Era

2026-05-19 · Emaad Khwaja, Chris Lettieri, Gerald Woo, Eden Belouadah 외 arxiv

We show that time series foundation models scale: a single training recipe produces reliable forecast-quality improvements from 4M to 2.5B parameters. We release Toto 2.0, a family of five open-weights forecasting models…

Time Series Forecasting

Don't Learn the Shape: Forecasting Periodic Time Series by Rank-1 Decomposition

2026-05-08 · Takato Honda arxiv

How few parameters do we really need to forecast a periodic time series? An hourly electricity series, reshaped as a 24-row matrix with one column per day, is approximately rank-1: a daily shape modulated by a daily leve…

GIFT: Generalizing Intent for Flexible Test-Time Rewards

2026-03-23 · Fin Amin, Nathaniel Dennler, Andreea Bobu arxiv

Robots learn reward functions from user demonstrations, but these rewards often fail to generalize to new environments. This failure occurs because learned rewards latch onto spurious correlations in training data rather…

Semantic Similarity

GIFT: Guided Fine-Tuning and Transfer for Enhancing Instruction-Tuned Language Models

2026-05-02 · Zhiwen Ruan, Yichao Du, Jianjie Zheng, Longyue Wang 외 arxiv

A promising paradigm for adapting instruction-tuned language models is to learn task-specific updates on a pretrained base model and subsequently merge them into the instruction-tuned model. However, existing approaches …