paper-with-me

Papers

DrFuse: Learning Disentangled Representation for Clinical Multi-Modal Fusion with Missing Modality and Modal Inconsistency

2024-03-10 · Wenfang Ya, Kejing Yin, William K. Cheung, Jia Liu, Jing Qin

The combination of electronic health records (EHR) and medical images is crucial for clinicians in making diagnoses and forecasting prognosis. Strategically fusing these two data modalities has great potential to improve the accuracy of machine learning models in clinical prediction tasks. However, the asynchronous and complementary nature of EHR and medical images presents unique challenges. Missing modalities due to clinical and administrative factors are inevitable in practice, and the significance of each data modality varies depending on the patient and the prediction target, resulting in inconsistent predictions and suboptimal model performance. To address these challenges, we propose DrFuse to achieve effective clinical multi-modal fusion. It tackles the missing modality issue by disentangling the features shared across modalities and those unique within each modality. Furthermore, we address the modal inconsistency issue via a disease-wise attention layer that produces the patient- and disease-wise weighting for each modality to make the final prediction. We validate the proposed method using real-world large-scale datasets, MIMIC-IV and MIMIC-CXR. Experimental results show that the proposed method significantly outperforms the state-of-the-art models. Our implementation is publicly available at https://github.com/dorothy-yao/drfuse.

📄 PDF Abstract BibTeX arXiv:2403.06197

Code (1)

dorothy-yao/drfuse 공식 구현 pytorch

Tasks

PredictionPrognosis

Similar Papers 제목 키워드 기반

MMDRFuse: Distilled Mini-Model with Dynamic Refresh for Multi-Modality Image Fusion

2024-08-28 · Yanglin Deng, Tianyang Xu, Chunyang Cheng, Xiao-Jun Wu 외

In recent years, Multi-Modality Image Fusion (MMIF) has been applied to many fields, which has attracted many scholars to endeavour to improve the fusion performance. However, the prevailing focus has predominantly been …

Pedestrian Detection

DiA-gnostic VLVAE: Disentangled Alignment-Constrained Vision Language Variational AutoEncoder for Robust Radiology Reporting with Missing Modalities

2025-11-08 · Nagur Shareef Shaik, Teja Krishna Cherukuri, Adnan Masood, Dong Hye Ye arxiv

The integration of medical images with clinical context is essential for generating accurate and clinically interpretable radiology reports. However, current automated methods often rely on resource-heavy Large Language …

Knowledge Graphs

CG-DMER: Hybrid Contrastive-Generative Framework for Disentangled Multimodal ECG Representation Learning

2026-02-24 · Ziwei Niu, Hao Sun, Shujun Bian, Xihong Yang 외 arxiv

Accurate interpretation of electrocardiogram (ECG) signals is crucial for diagnosing cardiovascular diseases. Recent multimodal approaches that integrate ECGs with accompanying clinical reports show strong potential, but…

Representation Learning

Multi-Modal Fusion for Sensorimotor Coordination in Steering Angle Prediction

2022-02-11 · Farzeen Munir, Shoaib Azam, Byung-Geun Lee, Moongu Jeon

Imitation learning is employed to learn sensorimotor coordination for steering angle prediction in an end-to-end fashion requires expert demonstrations. These expert demonstrations are paired with environmental perceptio…

Imitation Learning

Incomplete Modality Disentangled Representation for Ophthalmic Disease Grading and Diagnosis

2025-02-17 · Chengzhi Liu, Zile Huang, Zhe Chen, Feilong Tang 외

Ophthalmologists typically require multimodal data sources to improve diagnostic accuracy in clinical decisions. However, due to medical device shortages, low-quality data and data privacy concerns, missing data modaliti…

Diagnostic