paper-with-me

Papers

AutoRG-Brain: Grounded Report Generation for Brain MRI

2024-07-23 · Jiayu Lei, Xiaoman Zhang, Chaoyi Wu, Lisong Dai, Ya zhang, Yanyong Zhang, Yanfeng Wang, Weidi Xie, Yuehua Li

Radiologists are tasked with interpreting a large number of images in a daily base, with the responsibility of generating corresponding reports. This demanding workload elevates the risk of human error, potentially leading to treatment delays, increased healthcare costs, revenue loss, and operational inefficiencies. To address these challenges, we initiate a series of work on grounded Automatic Report Generation (AutoRG), starting from the brain MRI interpretation system, which supports the delineation of brain structures, the localization of anomalies, and the generation of well-organized findings. We make contributions from the following aspects, first, on dataset construction, we release a comprehensive dataset encompassing segmentation masks of anomaly regions and manually authored reports, termed as RadGenome-Brain MRI. This data resource is intended to catalyze ongoing research and development in the field of AI-assisted report generation systems. Second, on system design, we propose AutoRG-Brain, the first brain MRI report generation system with pixel-level grounded visual clues. Third, for evaluation, we conduct quantitative assessments and human evaluations of brain structure segmentation, anomaly localization, and report generation tasks to provide evidence of its reliability and accuracy. This system has been integrated into real clinical scenarios, where radiologists were instructed to write reports based on our generated findings and anomaly segmentation masks. The results demonstrate that our system enhances the report-writing skills of junior doctors, aligning their performance more closely with senior doctors, thereby boosting overall productivity.

📄 PDF Abstract BibTeX arXiv:2407.16684

Code (0)

등록된 구현이 없습니다.

Tasks

Anomaly LocalizationAnomaly Segmentation

Similar Papers 제목 키워드 기반

AGA3DNet: Anatomy-Guided Gaussian Priors with Multi-view xLSTM for 3D Brain MRI Subtype Classification

2026-05-08 · Peiyu Duan, Xueqi Guo, Sepehr Farhand, Mehmet Berk Sahin 외 arxiv

Accurate 3D brain MRI subtype classification benefits from both localized anatomical cues and long-range contextual reasoning. We present AGA3DNet, a report-grounded framework that incorporates brief anatomical phrases e…

3D Classification

Language Models for Automated Classification of Brain MRI Reports and Growth Chart Generation

2025-03-15 · Maryam Daniali, Shivaram Karandikar, Dabriel Zimmerman, J. Eric Schmitt 외

Clinically acquired brain MRIs and radiology reports are valuable but underutilized resources due to the challenges of manual analysis and data heterogeneity. We developed fine-tuned language models (LMs) to classify bra…

Benchmarking

Towards a Holistic Framework for Multimodal Large Language Models in Three-dimensional Brain CT Report Generation

2024-07-02 · Cheng-Yi Li, Kao-Jung Chang, Cheng-Fu Yang, Hsin-Yu Wu 외

Multi-modal large language models (MLLMs) have been given free rein to explore exciting medical applications with a primary focus on radiology report generation. Nevertheless, the preliminary success in 2D radiology capt…

AnatomyClinical KnowledgeDiagnosticMedical Report Generation+1

MEPNet: Medical Entity-balanced Prompting Network for Brain CT Report Generation

2025-03-22 · Xiaodan Zhang, Yanzhao Shi, Junzhong Ji, Chengxin Zheng 외

The automatic generation of brain CT reports has gained widespread attention, given its potential to assist radiologists in diagnosing cranial diseases. However, brain CT scans involve extensive medical entities, such as…

AnatomyLarge Language ModelText Generation

Multi-LLM Collaborative MRI Report Generation for Visual Instruction Tuning in Brain Oncology

2026-07-16 · Sinyoung Ra, Jonghun Kim, Hyunjin Park arxiv

Recent advances in large language models (LLMs) and their extension to vision-language models (VLMs) have made it easier to combine text and images for tasks such as report generation. Existing VLMs in medicine typically…

Visual Question Answering