paper-with-me

홈 › Papers

Learning ECG Image Representations via Dual Physiological-Aware Alignments

2026-04-02 · Hung Manh Pham, Jialu Tang, Aaqib Saeed, Dong Ma, Bin Zhu, Pan Zhou arxiv

Electrocardiograms (ECGs) are among the most widely used diagnostic tools for cardiovascular diseases, and a large amount of ECG data worldwide appears only in image form. However, most existing automated ECG analysis methods rely on access to raw signal recordings, limiting their applicability in real-world and resource-constrained settings. In this paper, we present ECG-Scan, a self-supervised framework for learning clinically generalized representations from ECG images through dual physiological-aware alignments: 1) Our approach optimizes image representation learning using multimodal contrastive alignment between image and gold-standard signal-text modalities. 2) We further integrate domain knowledge via soft-lead constraints, regularizing the reconstruction process and improving signal lead inter-consistency. Extensive benchmarking across multiple datasets and downstream tasks demonstrates that our image-based model achieves superior performance compared to existing image baselines and notably narrows the gap between ECG image and signal analysis. These results highlight the potential of self-supervised image modeling to unlock large-scale legacy ECG data and broaden access to automated cardiovascular diagnostics.

📄 PDF Abstract BibTeX arXiv:2604.01526

Code (0)

등록된 구현이 없습니다.

Tasks

Representation Learning

Similar Papers 제목 키워드 기반

Human Centered Non Intrusive Driver State Modeling Using Personalized Physiological Signals in Real World Automated Driving

2026-04-13 · David Puertas-Ramirez, Raul Fernandez-Matellan, David Martin Gomez, Jesus G. Boticario arxiv

In vehicles with partial or conditional driving automation (SAE Levels 2-3), the driver remains responsible for supervising the system and responding to take-over requests. Therefore, reliable driver monitoring is essent…

Similarity Reasoning and Filtration for Image-Text Matching

2021-01-05 · Haiwen Diao, Ying Zhang, Lin Ma, Huchuan Lu

Image-text matching plays a critical role in bridging the vision and language, and great progress has been made by exploiting the global alignment between image and sentence, or local alignments between regions and words…

Cross-Modal RetrievalImage RetrievalImage-text matchingSentence+2

Consensus-Aware Visual-Semantic Embedding for Image-Text Matching

2020-07-17 · ECCV 2020 8 · Haoran Wang, Ying Zhang, Zhong Ji, Yanwei Pang 외

Image-text matching plays a central role in bridging vision and language. Most existing approaches only rely on the image-text instance pair to learn their representations, thereby exploiting their matching relationships…

Image CaptioningImage-text matchingRetrievalText Matching+1

Context-Aware Interaction Network for Question Matching

2021-04-17 · EMNLP 2021 11 · Zhe Hu, Zuohui Fu, Yu Yin, Gerard de Melo

Impressive milestones have been achieved in text matching by adopting a cross-attention mechanism to capture pertinent semantic connections between two sentence representations. However, regular cross-attention focuses o…

SentenceText Matching

Cross-Temporal Attention Fusion (CTAF) for Multimodal Physiological Signals in Self-Supervised Learning

2026-02-02 · Arian Khorasani, Théophile Demazure arxiv

We study multimodal affect modeling when EEG and peripheral physiology are asynchronous, which most fusion methods ignore or handle with costly warping. We propose Cross-Temporal Attention Fusion (CTAF), a self-supervise…

Self-Supervised Learning