paper-with-me

홈 › Papers

Interpretable Medical Image Visual Question Answering via Multi-Modal Relationship Graph Learning

2023-02-19 · Xinyue Hu, Lin Gu, Kazuma Kobayashi, Qiyuan An, Qingyu Chen, Zhiyong Lu, Chang Su, Tatsuya Harada, Yingying Zhu

Medical visual question answering (VQA) aims to answer clinically relevant questions regarding input medical images. This technique has the potential to improve the efficiency of medical professionals while relieving the burden on the public health system, particularly in resource-poor countries. Existing medical VQA methods tend to encode medical images and learn the correspondence between visual features and questions without exploiting the spatial, semantic, or medical knowledge behind them. This is partially because of the small size of the current medical VQA dataset, which often includes simple questions. Therefore, we first collected a comprehensive and large-scale medical VQA dataset, focusing on chest X-ray images. The questions involved detailed relationships, such as disease names, locations, levels, and types in our dataset. Based on this dataset, we also propose a novel baseline method by constructing three different relationship graphs: spatial relationship, semantic relationship, and implicit relationship graphs on the image regions, questions, and semantic labels. The answer and graph reasoning paths are learned for different questions.

📄 PDF Abstract BibTeX arXiv:2302.09636

Code (0)

등록된 구현이 없습니다.

Tasks

Graph LearningMedical Visual Question AnsweringQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)

Similar Papers 제목 키워드 기반

Medico 2025: Visual Question Answering for Gastrointestinal Imaging

2025-08-14 · Sushant Gautam, Vajira Thambawita, Michael Riegler, Pål Halvorsen 외 arxiv

The Medico 2025 challenge addresses Visual Question Answering (VQA) for Gastrointestinal (GI) imaging, organized as part of the MediaEval task series. The challenge focuses on developing Explainable Artificial Intelligen…

Visual Question Answering

$M^3 QuestionIng$: Multi-modal Multi-span Medical Question Answering

2026-05-19 · Anisha Saha, Vaibhav Rathore, Abhisek Tiwari, Akash Ghosh 외 arxiv

The growing adoption of AI in healthcare, particularly in preventive care, highlights the critical need for accessibility and precision in Medical Question Answering (MedQA). In recent years, significant efforts have bee…

Question Answering

Structure Causal Models and LLMs Integration in Medical Visual Question Answering

2025-05-05 · Zibo Xu, Qiang Li, Weizhi Nie, Weijie Wang 외

Medical Visual Question Answering (MedVQA) aims to answer medical questions according to medical images. However, the complexity of medical data leads to confounders that are difficult to observe, so bias between images …

Causal InferenceMedical Visual Question AnsweringQuestion AnsweringVisual Question Answering

Localized Questions in Medical Visual Question Answering

2023-07-03 · Sergio Tascon-Morales, Pablo Márquez-Neila, Raphael Sznitman

Visual Question Answering (VQA) models aim to answer natural language questions about given images. Due to its ability to ask questions that differ from those used when training the model, medical VQA has received substa…

Medical Visual Question AnsweringQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)

An Interpretable Local Editing Model for Counterfactual Medical Image Generation

2026-02-28 · Hyungi Min, Taeseung You, Hangyeul Lee, Yeongjae Cho 외 arxiv

Counterfactual medical image generation have emerged as a critical tool for enhancing AI-driven systems in medical domain by answering "what-if" questions. However, existing approaches face two fundamental limitations: F…

Medical Image Generation