paper-with-me

Papers

Position: Empowering Time Series Reasoning with Multimodal LLMs

2025-02-03 · Yaxuan Kong, Yiyuan Yang, Shiyu Wang, Chenghao Liu, Yuxuan Liang, Ming Jin, Stefan Zohren, Dan Pei, Yan Liu, Qingsong Wen

Understanding time series data is crucial for multiple real-world applications. While large language models (LLMs) show promise in time series tasks, current approaches often rely on numerical data alone, overlooking the multimodal nature of time-dependent information, such as textual descriptions, visual data, and audio signals. Moreover, these methods underutilize LLMs' reasoning capabilities, limiting the analysis to surface-level interpretations instead of deeper temporal and multimodal reasoning. In this position paper, we argue that multimodal LLMs (MLLMs) can enable more powerful and flexible reasoning for time series analysis, enhancing decision-making and real-world applications. We call on researchers and practitioners to leverage this potential by developing strategies that prioritize trust, interpretability, and robust reasoning in MLLMs. Lastly, we highlight key research directions, including novel reasoning paradigms, architectural innovations, and domain-specific applications, to advance time series reasoning with MLLMs.

📄 PDF Abstract BibTeX arXiv:2502.01477

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingMultimodal ReasoningPositionTime SeriesTime Series Analysis

Similar Papers 제목 키워드 기반

Empowering VLMs for Few-Shot Multimodal Time Series Classification via Tailored Agentic Reasoning

2026-05-10 · Lin Li, Jiawei Huang, Qihao Quan, Dan Li 외 arxiv

In this paper, we propose the first VL$\underline{\textbf{M}}$ $\underline{\textbf{a}}$gentic $\underline{\textbf{r}}$easoning framework for few-$\underline{\textbf{s}}$hot multimodal $\underline{\textbf{T}}$ime $\underl…

Time Series Classification

GEM: Empowering MLLM for Grounded ECG Understanding with Time Series and Images

2025-03-08 · Xiang Lan, Feng Wu, Kai He, Qinghao Zhao 외

While recent multimodal large language models (MLLMs) have advanced automated ECG interpretation, they still face two key limitations: (1) insufficient multimodal synergy between time series signals and visual ECG repres…

cross-modal alignmentDiagnosticTime Series

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

2024-12-24 · Huanjin Yao, Jiaxing Huang, Wenhao Wu, Jingyi Zhang 외

In this work, we aim to develop an MLLM that understands and solves questions by learning to create each intermediate step of the reasoning involved till the final answer. To this end, we propose Collective Monte Carlo T…

Empowering Time Series Analysis with Large-Scale Multimodal Pretraining

2026-02-05 · Peng Chen, Siyuan Wang, Shiyan Hu, Xingjian Wu 외 arxiv

While existing time series foundation models primarily rely on large-scale unimodal pretraining, they lack complementary modalities to enhance time series understanding. Building multimodal foundation models is a natural…

Time Series ForecastingTime Series AnalysisAnomaly Detection

TS-Agent: Understanding and Reasoning Over Raw Time Series via Iterative Insight Gathering

2025-10-08 · Penghang Liu, Elizabeth Fons, Annita Vapsi, Mohsen Ghassemi 외 arxiv

Large language models (LLMs) exhibit strong symbolic and compositional reasoning, yet they struggle with time series question answering as the data is typically transformed into an LLM-compatible modality, e.g., serializ…

Question Answering