paper-with-me

홈 › Papers

On Creating a Causally Grounded Usable Rating Method for Assessing the Robustness of Foundation Models Supporting Time Series

2025-02-17 · Kausik Lakkaraju, Rachneet Kaur, Parisa Zehtabi, Sunandita Patra, Siva Likitha Valluru, Zhen Zeng, Biplav Srivastava, Marco Valtorta

Foundation Models (FMs) have improved time series forecasting in various sectors, such as finance, but their vulnerability to input disturbances can hinder their adoption by stakeholders, such as investors and analysts. To address this, we propose a causally grounded rating framework to study the robustness of Foundational Models for Time Series (FMTS) with respect to input perturbations. We evaluate our approach to the stock price prediction problem, a well-studied problem with easily accessible public data, evaluating six state-of-the-art (some multi-modal) FMTS across six prominent stocks spanning three industries. The ratings proposed by our framework effectively assess the robustness of FMTS and also offer actionable insights for model selection and deployment. Within the scope of our study, we find that (1) multi-modal FMTS exhibit better robustness and accuracy compared to their uni-modal versions and, (2) FMTS pre-trained on time series forecasting task exhibit better robustness and forecasting accuracy compared to general-purpose FMTS pre-trained across diverse settings. Further, to validate our framework's usability, we conduct a user study showcasing FMTS prediction errors along with our computed ratings. The study confirmed that our ratings reduced the difficulty for users in comparing the robustness of different systems.

📄 PDF Abstract BibTeX arXiv:2502.12226

Code (0)

등록된 구현이 없습니다.

Tasks

Model SelectionStock Price PredictionTime SeriesTime Series Forecasting

Similar Papers 제목 키워드 기반

LegalWorld: A Life-Cycle Interactive Environment for Legal Agents

2026-06-17 · Songhan Zuo, Shengbin Yue, Tao Chiang, Guanying Li 외 arxiv

Civil litigation is inherently a life-cycle process: what a lawyer drafts on day one constrains what unfolds at trial months later. Yet existing legal benchmarks evaluate isolated subtasks, and prior legal-agent simulato…

Parameter Estimation using Reinforcement Learning Causal Curiosity: Limits and Challenges

2025-05-13 · Miguel Arana-Catania, Weisi Guo

Causal understanding is important in many disciplines of science and engineering, where we seek to understand how different factors in the system causally affect an experiment or situation and pave a pathway towards crea…

Disentanglementparameter estimation

Do Large Language Models Reason Causally Like Us? Even Better?

2025-02-14 · Hanna M. Dettki, Brenden M. Lake, Charley M. Wu, Bob Rehder

Causal reasoning is a core component of intelligence. Large language models (LLMs) have shown impressive capabilities in generating human-like text, raising questions about whether their responses reflect true understand…

Decision Making

PrinciplismQA: A Philosophy-Grounded Approach to Assessing LLM-Human Clinical Medical Ethics Alignment

2025-08-07 · Chang Hong, Minghao Wu, Qingying Xiao, Yuchi Wang 외 arxiv

As medical LLMs transition to clinical deployment, assessing their ethical reasoning capability becomes critical. While achieving high accuracy on knowledge benchmarks, LLMs lack validated assessment for navigating ethic…

Causal Document-Grounded Dialogue Pre-training

2023-05-18 · Yingxiu Zhao, Bowen Yu, Haiyang Yu, Bowen Li 외

The goal of document-grounded dialogue (DocGD) is to generate a response by grounding the evidence in a supporting document in accordance with the dialogue context. This process involves four variables that are causally …