paper-with-me

홈 › Papers

Cardioformer: Advancing AI in ECG Analysis with Multi-Granularity Patching and ResNet

2025-05-08 · Md Kamrujjaman Mobin, Md Saiful Islam, Sadik Al Barid, Md Masum

Electrocardiogram (ECG) classification is crucial for automated cardiac disease diagnosis, yet existing methods often struggle to capture local morphological details and long-range temporal dependencies simultaneously. To address these challenges, we propose Cardioformer, a novel multi-granularity hybrid model that integrates cross-channel patching, hierarchical residual learning, and a two-stage self-attention mechanism. Cardioformer first encodes multi-scale token embeddings to capture fine-grained local features and global contextual information and then selectively fuses these representations through intra- and inter-granularity self-attention. Extensive evaluations on three benchmark ECG datasets under subject-independent settings demonstrate that model consistently outperforms four state-of-the-art baselines. Our Cardioformer model achieves the AUROC of 96.34$\pm$0.11, 89.99$\pm$0.12, and 95.59$\pm$1.66 in MIMIC-IV, PTB-XL and PTB dataset respectively outperforming PatchTST, Reformer, Transformer, and Medformer models. It also demonstrates strong cross-dataset generalization, achieving 49.18% AUROC on PTB and 68.41% on PTB-XL when trained on MIMIC-IV. These findings underscore the potential of Cardioformer to advance automated ECG analysis, paving the way for more accurate and robust cardiovascular disease diagnosis. We release the source code at https://github.com/KMobin555/Cardioformer.

📄 PDF Abstract BibTeX arXiv:2505.05538

Code (1)

kmobin555/cardioformer 공식 구현 pytorch

Tasks

ECG Classification

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Adafactor Adafactor is a stochastic optimization method based on Adam that reduces memory usage while retaining the empirical benefits of…
SentencePiece 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
1x1 Convolution A 1 x 1 Convolution is a convolution with some special properties in that it can be used for dimensionality reduction,…
Reversible Residual Block 설명 없음
Multi-Head Attention 설명 없음

Similar Papers 제목 키워드 기반

Medformer: A Multi-Granularity Patching Transformer for Medical Time-Series Classification

2024-05-24 · Yihe Wang, Nan Huang, Taida Li, Yujun Yan 외

Medical time series (MedTS) data, such as Electroencephalography (EEG) and Electrocardiography (ECG), play a crucial role in healthcare, such as diagnosing brain and heart diseases. Existing methods for MedTS classificat…

EEGElectrocardiography (ECG)Time SeriesTime Series Classification

Exploring the impact of spatiotemporal granularity on the demand prediction of dynamic ride-hailing

2022-03-19 · Kai Liu, Zhiju Chen, Toshiyuki Yamamoto, Liheng Tuo

Dynamic demand prediction is a key issue in ride-hailing dispatching. Many methods have been developed to improve the demand prediction accuracy of an increase in demand-responsive, ride-hailing transport services. Howev…

Prediction

Kairos: Toward Adaptive and Parameter-Efficient Time Series Foundation Models

2025-09-30 · Kun Feng, Shaocheng Lan, Yuchen Fang, Wenchao He 외 arxiv

Inherent temporal heterogeneity, such as varying sampling densities and periodic structures, has posed substantial challenges in zero-shot generalization for Time Series Foundation Models (TSFMs). Existing TSFMs predomin…

Zero-shot Generalization

Superscopes: Amplifying Internal Feature Representations for Language Model Interpretation

2025-03-03 · Jonathan Jacobi, Gal Niv

Understanding and interpreting the internal representations of large language models (LLMs) remains an open challenge. Patchscopes introduced a method for probing internal activations by patching them into new prompts, p…

Language ModelingLanguage Modelling

SAG-ViT: A Scale-Aware, High-Fidelity Patching Approach with Graph Attention for Vision Transformers

2024-11-14 · Shravan Venkatraman, Jaskaran Singh Walia, Joe Dhanith P R

Vision Transformers (ViTs) have redefined image classification by leveraging self-attention to capture complex patterns and long-range dependencies between image patches. However, a key challenge for ViTs is efficiently …

Graph Attentionimage-classificationImage Classification