paper-with-me

홈 › Papers

FODA-PG for Enhanced Medical Imaging Narrative Generation: Adaptive Differentiation of Normal and Abnormal Attributes

2024-09-06 · Kai Shu, Yuzhuo Jia, Ziyang Zhang, Jiechao Gao

Automatic Medical Imaging Narrative generation aims to alleviate the workload of radiologists by producing accurate clinical descriptions directly from radiological images. However, the subtle visual nuances and domain-specific terminology in medical images pose significant challenges compared to generic image captioning tasks. Existing approaches often neglect the vital distinction between normal and abnormal findings, leading to suboptimal performance. In this work, we propose FODA-PG, a novel Fine-grained Organ-Disease Adaptive Partitioning Graph framework that addresses these limitations through domain-adaptive learning. FODA-PG constructs a granular graphical representation of radiological findings by separating disease-related attributes into distinct "disease-specific" and "disease-free" categories based on their clinical significance and location. This adaptive partitioning enables our model to capture the nuanced differences between normal and pathological states, mitigating the impact of data biases. By integrating this fine-grained semantic knowledge into a powerful transformer-based architecture and providing rigorous mathematical justifications for its effectiveness, FODA-PG generates precise and clinically coherent reports with enhanced generalization capabilities. Extensive experiments on the IU-Xray and MIMIC-CXR benchmarks demonstrate the superiority of our approach over state-of-the-art methods, highlighting the importance of domain adaptation in medical report generation.

📄 PDF Abstract BibTeX arXiv:2409.03947

Code (0)

등록된 구현이 없습니다.

Tasks

Domain AdaptationImage CaptioningMedical Report Generation

Similar Papers 제목 키워드 기반

MedicalNarratives: Connecting Medical Vision and Language with Localized Narratives

2025-01-07 · Wisdom O. Ikezogwo, Kevin Zhang, Mehmet Saygin Seyfioglu, Fatemeh Ghezloo 외

We propose MedicalNarratives, a dataset curated from medical pedagogical videos similar in nature to data collected in Think-Aloud studies and inspired by Localized Narratives, which collects grounded image-text data by …

Articles

Infogen: Generating Complex Statistical Infographics from Documents

2025-07-26 · Akash Ghosh, Aparna Garimella, Pritika Ramu, Sambaran Bandyopadhyay 외 arxiv

Statistical infographics are powerful tools that simplify complex data into visually engaging and easy-to-understand formats. Despite advancements in AI, particularly with LLMs, existing efforts have been limited to gene…

A Novel Corpus of Annotated Medical Imaging Reports and Information Extraction Results Using BERT-based Language Models

2024-03-27 · Namu Park, Kevin Lybarger, Giridhar Kaushik Ramachandran, Spencer Lewis 외

Medical imaging is critical to the diagnosis, surveillance, and treatment of many health conditions, including oncological, neurological, cardiovascular, and musculoskeletal disorders, among others. Radiologists interpre…

Anatomy

ProbMed: A Probabilistic Framework for Medical Multimodal Binding

2025-09-30 · Yuan Gao, Sangwook Kim, Jianzhong You, Chris McIntosh arxiv

Medical decision-making requires integrating diverse medical information, from imaging to clinical narratives. These medical modalities are often acquired in a many-to-many manner. However, current medical vision-languag…

Contrastive Learning

Physics-driven Synthetic Data Learning for Biomedical Magnetic Resonance

2022-03-21 · Qinqin Yang, Zi Wang, Kunyuan Guo, Congbo Cai 외

Deep learning has innovated the field of computational imaging. One of its bottlenecks is unavailable or insufficient training data. This article reviews an emerging paradigm, imaging physics-based data synthesis (IPADS)…

Deep Learning