paper-with-me

Papers

Context Clues: Evaluating Long Context Models for Clinical Prediction Tasks on EHRs

2024-12-09 · Michael Wornow, Suhana Bedi, Miguel Angel Fuentes Hernandez, Ethan Steinberg, Jason Alan Fries, Christopher Re, Sanmi Koyejo, Nigam H. Shah

Foundation Models (FMs) trained on Electronic Health Records (EHRs) have achieved state-of-the-art results on numerous clinical prediction tasks. However, most existing EHR FMs have context windows of <1k tokens. This prevents them from modeling full patient EHRs which can exceed 10k's of events. Recent advancements in subquadratic long-context architectures (e.g., Mamba) offer a promising solution. However, their application to EHR data has not been well-studied. We address this gap by presenting the first systematic evaluation of the effect of context length on modeling EHR data. We find that longer context models improve predictive performance -- our Mamba-based model surpasses the prior state-of-the-art on 9/14 tasks on the EHRSHOT prediction benchmark. For clinical applications, however, model performance alone is insufficient -- robustness to the unique properties of EHR is crucial. Thus, we also evaluate models across three previously underexplored properties of EHR data: (1) the prevalence of "copy-forwarded" diagnoses which creates artificial repetition of tokens within EHR sequences; (2) the irregular time intervals between EHR events which can lead to a wide range of timespans within a context window; and (3) the natural increase in disease complexity over time which makes later tokens in the EHR harder to predict than earlier ones. Stratifying our EHRSHOT results, we find that higher levels of each property correlate negatively with model performance, but that longer context models are more robust to more extreme levels of these properties. Our work highlights the potential for using long-context architectures to model EHR data, and offers a case study for identifying new challenges in modeling sequential data motivated by domains outside of natural language. We release our models and code at: https://github.com/som-shahlab/long_context_clues

📄 PDF Abstract BibTeX arXiv:2412.16178

Code (2)

som-shahlab/long_context_clues 공식 구현 jax
reAIM-Lab/ehr_foundation_model_benchmark

Tasks

Mamba

Similar Papers 제목 키워드 기반

Video Detective: Seek Critical Clues Recurrently to Answer Question from Long Videos

2025-12-19 · Henghui Du, Chunjie Zhang, Xi Chen, Chang Zhou 외 arxiv

Long Video Question-Answering (LVQA) presents a significant challenge for Multi-modal Large Language Models (MLLMs) due to immense context and overloaded information, which could also lead to prohibitive memory consumpti…

Evaluating Retrieval-Augmented Generation vs. Long-Context Input for Clinical Reasoning over EHRs

2025-08-20 · Skatje Myers, Dmitriy Dligach, Timothy A. Miller, Samantha Barr 외 arxiv

Objective: To evaluate whether retrieval-augmented generation (RAG) can serve as an efficient alternative to long-context prompting for clinical reasoning over electronic health records (EHRs). Methods: We defined three …

Text Retrieval

Story Ending Generation with Incremental Encoding and Commonsense Knowledge

2018-08-30 · Jian Guan, Yansen Wang, Minlie Huang

Generating a reasonable ending for a given story context, i.e., story ending generation, is a strong indication of story comprehension. This task requires not only to understand the context clues which play an important …

Image-guided Story Ending Generation

Video-based Person Re-identification with Accumulative Motion Context

2017-06-13 · Hao liu, Zequn Jie, Karlekar Jayashree, Meibin Qi 외

Video based person re-identification plays a central role in realistic security and video surveillance. In this paper we propose a novel Accumulative Motion Context (AMOC) network for addressing this important problem, w…

Person Re-IdentificationVideo-Based Person Re-Identification

Video-based Person Re-identification with Accumulative Motion Context

2017-01-01 · Hao Liu, Zequn Jie, Karlekar Jayashree, Meibin Qi 외

Video based person re-identification plays a central role in realistic security and video surveillance. In this paper we propose a novel Accumulative Motion Context (AMOC) network for addressing this important problem, w…

Person Re-IdentificationVideo-Based Person Re-Identification