paper-with-me

홈 › Papers

Using Multi-modal Data for Improving Generalizability and Explainability of Disease Classification in Radiology

2022-07-29 · Pranav Agnihotri, Sara Ketabi, Khashayar, Namdar, Farzad Khalvati

Traditional datasets for the radiological diagnosis tend to only provide the radiology image alongside the radiology report. However, radiology reading as performed by radiologists is a complex process, and information such as the radiologist's eye-fixations over the course of the reading has the potential to be an invaluable data source to learn from. Nonetheless, the collection of such data is expensive and time-consuming. This leads to the question of whether such data is worth the investment to collect. This paper utilizes the recently published Eye-Gaze dataset to perform an exhaustive study on the impact on performance and explainability of deep learning (DL) classification in the face of varying levels of input features, namely: radiology images, radiology report text, and radiologist eye-gaze data. We find that the best classification performance of X-ray images is achieved with a combination of radiology report free-text and radiology image, with the eye-gaze data providing no performance boost. Nonetheless, eye-gaze data serving as secondary ground truth alongside the class label results in highly explainable models that generate better attention maps compared to models trained to do classification and attention map generation without eye-gaze data.

📄 PDF Abstract BibTeX arXiv:2207.14781

Code (0)

등록된 구현이 없습니다.

Tasks

Explainable Models

Similar Papers 제목 키워드 기반

MedViLaM: A multimodal large language model with advanced generalizability and explainability for medical data understanding and generation

2024-09-29 · Lijian Xu, Hao Sun, Ziyu Ni, Hongsheng Li 외

Medicine is inherently multimodal and multitask, with diverse data modalities spanning text, imaging. However, most models in medical field are unimodal single tasks and lack good generalizability and explainability. In …

Language ModelingLanguage ModellingLarge Language ModelMultimodal Large Language Model+3

VisionFM: a Multi-Modal Multi-Task Vision Foundation Model for Generalist Ophthalmic Artificial Intelligence

2023-10-08 · Jianing Qiu, Jian Wu, Hao Wei, Peilun Shi 외

We present VisionFM, a foundation model pre-trained with 3.4 million ophthalmic images from 560,457 individuals, covering a broad range of ophthalmic diseases, modalities, imaging devices, and demography. After pre-train…

Disease PredictionPrognosisRepresentation Learning

Framework for developing and evaluating ethical collaboration between expert and machine

2024-11-17 · Ayan Banerjee, Payal Kamboj, Sandeep Gupta

Precision medicine is a promising approach for accessible disease diagnosis and personalized intervention planning in high-mortality diseases such as coronary artery disease (CAD), drug-resistant epilepsy (DRE), and chro…

Decision MakingManagement

DDCoT: Duty-Distinct Chain-of-Thought Prompting for Multimodal Reasoning in Language Models

2023-10-25 · NeurIPS 2023 11

A long-standing goal of AI systems is to perform complex multimodal reasoning like humans. Recently, large language models (LLMs) have made remarkable strides in such multi-step reasoning on the language modality solely …

Multimodal Reasoning

Anatomy-Aware Conditional Image-Text Retrieval

2025-03-10 · Meng Zheng, Jiajin Zhang, Benjamin Planche, Zhongpai Gao 외

Image-Text Retrieval (ITR) finds broad applications in healthcare, aiding clinicians and radiologists by automatically retrieving relevant patient cases in the database given the query image and/or report, for more effic…

AnatomyContrastive LearningImage-text RetrievalRetrieval+1