paper-with-me

Papers

Context-driven Missing-Modality Learning for Robust Medical Diagnosis with Image-Tabular Data

2026-05-25 · Tianling Liu, Lequan Yu, Tong Han, Liang Wan arxiv

While multimodal data integrating diverse imaging and clinical tabular records is crucial for accurate medical diagnosis, the arbitrary absence of specific modalities is prevalent in clinical practice, severely degrading the performance of multimodal models. Existing methods either discard missing modalities, leading to information loss, or struggle to synthesize them without capturing complex inter-modal dependencies. To address these limitations, we propose a novel Context-driven Missing-Modality Learning (CMML) framework, which sequentially performs modality synthesis and semantic alignment to achieve robust diagnosis under arbitrary missing conditions. Specifically, we design a Cascade Residual Transformer-based Autoencoder (CRTA) that leverages learnable context tokens acting as dataset-level semantic prior to capture inter-modal dependencies and synthesize key missing representations. These representations are further enriched by modality-specific memory banks. To resolve the discrepancy between original available and synthesized representations, we transform the learned context tokens into instance-adaptive semantic references by infusing multimodal representations from the CRTA's outputs. This reference guides the alignment of heterogeneous modality representations into a unified space, where class-aware contrastive refinement is finally applied to explore discriminative diagnostic cues. Extensive evaluations on skin lesion (Derm7pt), ocular disease (ODIR), and meningioma (MEN) datasets demonstrate that CMML significantly outperforms state-of-the-art (SOTA) methods, yielding AVG AUC improvements of 1.26%, 0.97%, and 1.32%, respectively.

📄 PDF Abstract BibTeX arXiv:2605.25968

Code (0)

등록된 구현이 없습니다.

Tasks

Medical Diagnosis

Similar Papers 제목 키워드 기반

Missing-modality Enabled Multi-modal Fusion Architecture for Medical Data

2023-09-27 · Muyu Wang, Shiyu Fan, Yichen Li, Hui Chen

Fusing multi-modal data can improve the performance of deep learning models. However, missing modalities are common for medical data due to patients' specificity, which is detrimental to the performance of multi-modal mo…

Specificity

Med-K2N: Flexible K-to-N Modality Translation for Medical Image Synthesis

2025-10-03 · Feng Yuan, Yifan Gao, Yuehua Ye, Haoyue Li 외 arxiv

Cross-modal medical image synthesis research focuses on reconstructing missing imaging modalities from available ones to support clinical diagnosis. Driven by clinical necessities for flexible modality reconstruction, we…

How can Deep Learning Retrieve the Write-Missing Additional Diagnosis from Chinese Electronic Medical Record For DRG

2023-03-28 · Shaohui Liu, Xien Liu, Ji Wu

The purpose of write-missing diagnosis detection is to find diseases that have been clearly diagnosed from medical records but are missed in the discharge diagnosis. Unlike the definition of missed diagnosis, the write-m…

Graph Convolutional Networks for Multi-modality Medical Imaging: Methods, Architectures, and Clinical Applications

2022-02-17 · Kexin Ding, Mu Zhou, Zichen Wang, Qiao Liu 외

Image-based characterization and disease understanding involve integrative analysis of morphological, spatial, and topological information across biological scales. The development of graph convolutional networks (GCNs) …

Medical Image Analysis

MedMIX: Modality-Internal Expert Fusion for Multimodal Medical Diagnosis

2026-05-15 · Seungik Cho, Anqi Li, Wei Qiu arxiv

Multimodal clinical prediction faces three challenges: multiple foundation models (FMs) with complementary strengths per modality, pervasive missing modalities at training and test time, and sample-specific variation in …

Medical Diagnosis