paper-with-me

홈 › Papers

Revisiting Integration of Image and Metadata for DICOM Series Classification: Cross-Attention and Dictionary Learning

2026-02-27 · Tuan Truong, Melanie Dohmen, Sara Lorio, Matthias Lenga arxiv

Automated identification of DICOM image series is essential for large-scale medical image analysis, quality control, protocol harmonization, and reliable downstream processing. However, DICOM series classification remains challenging due to heterogeneous slice content, variable series length, and entirely missing, incomplete or inconsistent DICOM metadata. We propose an end-to-end multimodal framework for DICOM series classification that jointly models image content and acquisition metadata while explicitly accounting for all these challenges. (i) Images and metadata are encoded with modality-aware modules and fused using a bi-directional cross-modal attention mechanism. (ii) Metadata is processed by a sparse, missingness-aware encoder based on learnable feature dictionaries and value-conditioned modulation. By design, the approach does not require any form of imputation. (iii) Variability in series length and image data dimensions is handled via a 2.5D visual encoder and attention operating on equidistantly sampled slices. We evaluate the proposed approach on the publicly available Duke Liver MRI dataset and a large multi-institutional in-house cohort, assessing both in-domain performance and out-of-domain generalization. Across all evaluation settings, the proposed method consistently outperforms relevant image only, metadata-only and multimodal 2D/3D baselines. The results demonstrate that explicitly modeling metadata sparsity and cross-modal interactions improves robustness for DICOM series classification.

📄 PDF Abstract BibTeX arXiv:2602.23833

Code (0)

등록된 구현이 없습니다.

Tasks

Domain Generalization

Similar Papers 제목 키워드 기반

Automatic classification of prostate MR series type using image content and metadata

2024-04-16 · Deepa Krishnaswamy, Bálint Kovács, Stefan Denner, Steve Pieper 외

With the wealth of medical image data, efficient curation is essential. Assigning the sequence type to magnetic resonance images is necessary for scientific studies and artificial intelligence-based analysis. However, in…

Weakly Supervised Context Encoder using DICOM metadata in Ultrasound Imaging

2020-03-20 · Szu-Yeu Hu, Shuhang Wang, Wei-Hung Weng, JingChao Wang 외

Modern deep learning algorithms geared towards clinical adaption rely on a significant amount of high fidelity labeled data. Low-resource settings pose challenges like acquiring high fidelity data and becomes the bottlen…

Medical Image De-Identification Resources: Synthetic DICOM Data and Tools for Validation

2025-08-03 · Michael W. Rutherford, Tracy Nolan, Linmin Pei, Ulrike Wagner 외 arxiv

Medical imaging research increasingly depends on large-scale data sharing to promote reproducibility and train Artificial Intelligence (AI) models. Ensuring patient privacy remains a significant challenge for open-access…

DICODerma: A practical approach for metadata management of images in dermatology

2021-02-17 · Bell Raj Eapen, Feroze Kaliyadan, Ashique Karalikkattil T

Clinical images are vital for diagnosing and monitoring skin diseases, and their importance has increased with the growing popularity of machine learning. Lack of standards has stifled innovation in dermatological imagin…

BIG-bench Machine LearningManagementTAG

Multi-Modal Dataset Creation for Federated Learning with DICOM Structured Reports

2024-07-12 · Malte Tölle, Lukas Burger, Halvar Kelm, Florian André 외

Purpose: Federated training is often hindered by heterogeneous datasets due to divergent data storage options, inconsistent naming schemes, varied annotation procedures, and disparities in label quality. This is particul…

Data IntegrationFederated Learning