paper-with-me

홈 › Papers

MEETI: A Multimodal ECG Dataset from MIMIC-IV-ECG with Signals, Images, Features and Interpretations

2025-07-21 · Deyun Zhang, Xiang Lan, Shijia Geng, Qinghao Zhao, Sumei Fan, Mengling Feng, Shenda Hong arxiv

Electrocardiogram (ECG) plays a foundational role in modern cardiovascular care, enabling non-invasive diagnosis of arrhythmias, myocardial ischemia, and conduction disorders. While machine learning has achieved expert-level performance in ECG interpretation, the development of clinically deployable multimodal AI systems remains constrained, primarily due to the lack of publicly available datasets that simultaneously incorporate raw signals, diagnostic images, and interpretation text. Most existing ECG datasets provide only single-modality data or, at most, dual modalities, making it difficult to build models that can understand and integrate diverse ECG information in real-world settings. To address this gap, we introduce MEETI (MIMIC-IV-Ext ECG-Text-Image), the first large-scale ECG dataset that synchronizes raw waveform data, high-resolution plotted images, and detailed textual interpretations generated by large language models. In addition, MEETI includes beat-level quantitative ECG parameters extracted from each lead, offering structured parameters that support fine-grained analysis and model interpretability. Each MEETI record is aligned across four components: (1) the raw ECG waveform, (2) the corresponding plotted image, (3) extracted feature parameters, and (4) detailed interpretation text. This alignment is achieved using consistent, unique identifiers. This unified structure supports transformer-based multimodal learning and supports fine-grained, interpretable reasoning about cardiac health. By bridging the gap between traditional signal analysis, image-based interpretation, and language-driven understanding, MEETI established a robust foundation for the next generation of explainable, multimodal cardiovascular AI. It offers the research community a comprehensive benchmark for developing and evaluating ECG-based AI systems.

📄 PDF Abstract BibTeX arXiv:2507.15255

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

You need to MIMIC to get FAME: Solving Meeting Transcript Scarcity with a Multi-Agent Conversations

2025-02-18 · Frederic Kirstein, Muneeb Khan, Jan Philip Wahle, Terry Ruas 외

Meeting summarization suffers from limited high-quality data, mainly due to privacy restrictions and expensive collection processes. We address this gap with FAME, a dataset of 500 meetings in English and 300 in German p…

Language ModelingLanguage ModellingLarge Language ModelMeeting Summarization

MIMIC: A Generative Multimodal Foundation Model for Biomolecules

2026-04-27 · Siavash Golkar, Jake Kovalic, Irina Espejo Morales, Samuel Sledzieski 외 arxiv

Biological function emerges from coupled constraints across sequence, structure, regulation, evolution, and cellular context, yet most foundation models in biology are trained within one modality or for a fixed forward t…

Representation Learning

Learning Dynamic Representations and Policies from Multimodal Clinical Time-Series with Informative Missingness

2026-04-23 · Zihan Liang, Ziwen Pan, Ruoxuan Xiong arxiv

Multimodal clinical records contain structured measurements and clinical notes recorded over time, offering rich temporal information about the evolution of patient health. Yet these observations are sparse, and whether …

Representation LearningMortality Prediction

Introducing Multimodal Paradigm for Learning Sleep Staging PSG via General-Purpose Model

2025-09-26 · Jianheng Zhou, Chenyu Liu, Jinan Zhou, Yi Ding 외 arxiv

Sleep staging is essential for diagnosing sleep disorders and assessing neurological health. Existing automatic methods typically extract features from complex polysomnography (PSG) signals and train domain-specific mode…

Anchoring Emotions in Text: Robust Multimodal Fusion for Mimicry Intensity Estimation

2026-03-16 · Lingsi Zhu, Yuefeng Zou, Yunxiang Zhang, Naixiang Zheng 외 arxiv

Estimating Emotional Mimicry Intensity (EMI) in naturalistic environments is a critical yet challenging task in affective computing. The primary difficulty lies in effectively modeling the complex, nonlinear temporal dyn…