paper-with-me

홈 › Papers

A Self-Supervised Model for Multi-modal Stroke Risk Prediction

2024-11-14 · Camille Delgrange, Olga Demler, Samia Mora, Bjoern Menze, Ezequiel de la Rosa, Neda Davoudi

Predicting stroke risk is a complex challenge that can be enhanced by integrating diverse clinically available data modalities. This study introduces a self-supervised multimodal framework that combines 3D brain imaging, clinical data, and image-derived features to improve stroke risk prediction prior to onset. By leveraging large unannotated clinical datasets, the framework captures complementary and synergistic information across image and tabular data modalities. Our approach is based on a contrastive learning framework that couples contrastive language-image pretraining with an image-tabular matching module, to better align multimodal data representations in a shared latent space. The model is trained on the UK Biobank, which includes structural brain MRI and clinical data. We benchmark its performance against state-of-the-art unimodal and multimodal methods using tabular, image, and image-tabular combinations under diverse frozen and trainable model settings. The proposed model outperformed self-supervised tabular (image) methods by 2.6% (2.6%) in ROC-AUC and by 3.3% (5.6%) in balanced accuracy. Additionally, it showed a 7.6% increase in balanced accuracy compared to the best multimodal supervised model. Through interpretable tools, our approach demonstrated better integration of tabular and image data, providing richer and more aligned embeddings. Gradient-weighted Class Activation Mapping heatmaps further revealed activated brain regions commonly associated in the literature with brain aging, stroke risk, and clinical outcomes. This robust self-supervised multimodal framework surpasses state-of-the-art methods for stroke risk prediction and offers a strong foundation for future studies integrating diverse data modalities to advance clinical predictive modelling.

📄 PDF Abstract BibTeX arXiv:2411.09822

Code (1)

CamilleDelgrange/SSMSRPM 공식 구현 pytorch

Tasks

Contrastive Learning

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음
ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Multimodal Deep Learning for Stroke Prediction and Detection using Retinal Imaging and Clinical Data

2025-05-05 · Saeed Shurrab, Aadim Nepal, Terrence J. Lee-St. John, Nicola G. Ghazi 외

Stroke is a major public health problem, affecting millions worldwide. Deep learning has recently demonstrated promise for enhancing the diagnosis and risk prediction of stroke. However, existing methods rely on costly m…

Multimodal Deep LearningSelf-Supervised Learning

Predicting Stroke through Retinal Graphs and Multimodal Self-supervised Learning

2024-11-08 · Yuqing Huang, Bastian Wittmann, Olga Demler, Bjoern Menze 외

Early identification of stroke is crucial for intervention, requiring reliable models. We proposed an efficient retinal image representation together with clinical information to capture a comprehensive overview of cardi…

Contrastive LearningSelf-Supervised LearningTransfer Learning

Whole-body Representation Learning For Competing Preclinical Disease Risk Assessment

2025-08-04 · Dmitrii Seletkov, Sophie Starck, Ayhan Can Erdur, Yundi Zhang 외 arxiv

Reliable preclinical disease risk assessment is essential to move public healthcare from reactive treatment to proactive identification and prevention. However, image-based risk prediction algorithms often consider one c…

Representation Learning

FAST-CAD: A Fairness-Aware Framework for Non-Contact Stroke Diagnosis

2025-11-12 · Tommy Sha, Zhan Cheng, Haotian Zhai, Xuwei Ding 외 arxiv

Stroke is an acute cerebrovascular disease, and timely diagnosis significantly improves patient survival. However, existing automated diagnosis methods suffer from fairness issues across demographic groups, potentially e…

Domain Adaptation

XMP-Font: Self-Supervised Cross-Modality Pre-training for Few-Shot Font Generation

2022-04-11 · CVPR 2022 1 · Wei Liu, Fangyue Liu, Fei Ding, Qian He 외

Generating a new font library is a very labor-intensive and time-consuming job for glyph-rich scripts. Few-shot font generation is thus required, as it requires only a few glyph references without fine-tuning during test…

DisentanglementFont Generation