paper-with-me

홈 › Papers

Benchmarking Robustness of Contrastive Learning Models for Medical Image-Report Retrieval

2025-01-15 · Demetrio Deanda, Yuktha Priya Masupalli, Jeong Yang, Young Lee, Zechun Cao, Gongbo Liang

Medical images and reports offer invaluable insights into patient health. The heterogeneity and complexity of these data hinder effective analysis. To bridge this gap, we investigate contrastive learning models for cross-domain retrieval, which associates medical images with their corresponding clinical reports. This study benchmarks the robustness of four state-of-the-art contrastive learning models: CLIP, CXR-RePaiR, MedCLIP, and CXR-CLIP. We introduce an occlusion retrieval task to evaluate model performance under varying levels of image corruption. Our findings reveal that all evaluated models are highly sensitive to out-of-distribution data, as evidenced by the proportional decrease in performance with increasing occlusion levels. While MedCLIP exhibits slightly more robustness, its overall performance remains significantly behind CXR-CLIP and CXR-RePaiR. CLIP, trained on a general-purpose dataset, struggles with medical image-report retrieval, highlighting the importance of domain-specific training data. The evaluation of this work suggests that more effort needs to be spent on improving the robustness of these models. By addressing these limitations, we can develop more reliable cross-domain retrieval models for medical applications.

📄 PDF Abstract BibTeX arXiv:2501.09134

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingContrastive LearningRetrieval

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음
CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…

Similar Papers 제목 키워드 기반

CXPMRG-Bench: Pre-training and Benchmarking for X-ray Medical Report Generation on CheXpert Plus Dataset

2024-10-01 · CVPR 2025 1 · Xiao Wang, Fuling Wang, Yuehang Li, Qingchuan Ma 외

X-ray image-based medical report generation (MRG) is a pivotal area in artificial intelligence which can significantly reduce diagnostic burdens and patient wait times. Despite significant progress, we believe that the t…

BenchmarkingContrastive LearningDiagnosticMamba+1

Benchmarking Vision-Language Contrastive Methods for Medical Representation Learning

2024-06-11 · Shuvendu Roy, Yasaman Parhizkar, Franklin Ogidi, Vahid Reza Khazaie 외

We perform a comprehensive benchmarking of contrastive frameworks for learning multimodal representations in the medical domain. Through this study, we aim to answer the following research questions: (i) How transferable…

BenchmarkingContrastive LearningImage RetrievalImage to text+4

Medical Report Generation based on Segment-Enhanced Contrastive Representation Learning

2023-12-26 · Ruoqing Zhao, Xi Wang, Hongliang Dai, Pan Gao 외

Automated radiology report generation has the potential to improve radiology reporting and alleviate the workload of radiologists. However, the medical report generation task poses unique challenges due to the limited av…

Contrastive LearningImage SegmentationMedical Image SegmentationMedical Report Generation+2

Representative Image Feature Extraction via Contrastive Learning Pretraining for Chest X-ray Report Generation

2022-09-04 · Yu-Jen Chen, Wei-Hsiang Shen, Hao-Wei Chung, Ching-Hao Chiu 외

Medical report generation is a challenging task since it is time-consuming and requires expertise from experienced radiologists. The goal of medical report generation is to accurately capture and describe the image findi…

Contrastive LearningMedical Report GenerationSegmentation

Mining Gaze for Contrastive Learning toward Computer-Assisted Diagnosis

2023-12-11 · Zihao Zhao, Sheng Wang, Qian Wang, Dinggang Shen

Obtaining large-scale radiology reports can be difficult for medical images due to various reasons, limiting the effectiveness of contrastive pre-training in the medical image domain and underscoring the need for alterna…

Contrastive LearningSemantic SimilaritySemantic Textual Similarity