paper-with-me

홈 › Papers

Self-supervised learning of imaging and clinical signatures using a multimodal joint-embedding predictive architecture

2025-09-18 · Thomas Z. Li, Aravind R. Krishnan, Lianrui Zuo, John M. Still, Kim L. Sandler, Fabien Maldonado, Thomas A. Lasko, Bennett A. Landman arxiv

The development of multimodal models for pulmonary nodule diagnosis is limited by the scarcity of labeled data and the tendency for these models to overfit on the training distribution. In this work, we leverage self-supervised learning from longitudinal and multimodal archives to address these challenges. We curate an unlabeled set of patients with CT scans and linked electronic health records from our home institution to power joint embedding predictive architecture (JEPA) pretraining. After supervised finetuning, we show that our approach outperforms an unregularized multimodal model and imaging-only model in an internal cohort (ours: 0.91, multimodal: 0.88, imaging-only: 0.73 AUC), but underperforms in an external cohort (ours: 0.72, imaging-only: 0.75 AUC). We develop a synthetic environment that characterizes the context in which JEPA may underperform. This work innovates an approach that leverages unlabeled multimodal medical archives to improve predictive models and demonstrates its advantages and limitations in pulmonary nodule diagnosis.

📄 PDF Abstract BibTeX arXiv:2509.15470

Code (0)

등록된 구현이 없습니다.

Tasks

Self-Supervised Learning

Similar Papers 제목 키워드 기반

Longitudinal Multimodal Transformer Integrating Imaging and Latent Clinical Signatures From Routine EHRs for Pulmonary Nodule Classification

2023-04-06 · Thomas Z. Li, John M. Still, Kaiwen Xu, Ho Hin Lee 외

The accuracy of predictive models for solitary pulmonary nodule (SPN) diagnosis can be greatly increased by incorporating repeat imaging and medical context, such as electronic health records (EHRs). However, clinically …

Computed Tomography (CT)DiagnosticDisentanglement

Multimodal Foundation Models for Early Disease Detection

2025-10-02 · Md Talha Mohsin, Ismail Abdulrashid arxiv

Healthcare data now span EHRs, medical imaging, genomics, and wearable sensors, but most diagnostic models still process these modalities in isolation. This limits their ability to capture early, cross-modal disease sign…

A Self-Supervised Model for Multi-modal Stroke Risk Prediction

2024-11-14 · Camille Delgrange, Olga Demler, Samia Mora, Bjoern Menze 외

Predicting stroke risk is a complex challenge that can be enhanced by integrating diverse clinically available data modalities. This study introduces a self-supervised multimodal framework that combines 3D brain imaging,…

Contrastive Learning

Self-supervised multimodal neuroimaging yields predictive representations for a spectrum of Alzheimer's phenotypes

2022-09-07 · Alex Fedorov, Eloy Geenjaar, Lei Wu, Tristan Sylvain 외

Recent neuroimaging studies that focus on predicting brain disorders via modern machine learning approaches commonly include a single modality and rely on supervised over-parameterized models.However, a single modality p…

DiagnosticSelf-Supervised Learning

PRISM: A Framework Harnessing Unsupervised Visual Representations and Textual Prompts for Explainable MACE Survival Prediction from Cardiac Cine MRI

2025-08-26 · Haoyang Su, Jin-Yi Xiang, Shaohao Rui, Yifan Gao 외 arxiv

Accurate prediction of major adverse cardiac events (MACE) remains a central challenge in cardiovascular prognosis. We present PRISM (Prompt-guided Representation Integration for Survival Modeling), a self-supervised fra…