paper-with-me

Papers

Ultrasound Report Generation with Cross-Modality Feature Alignment via Unsupervised Guidance

2024-06-02 · Jun Li, Tongkun Su, Baoliang Zhao, Faqin Lv, Qiong Wang, Nassir Navab, Ying Hu, Zhongliang Jiang

Automatic report generation has arisen as a significant research area in computer-aided diagnosis, aiming to alleviate the burden on clinicians by generating reports automatically based on medical images. In this work, we propose a novel framework for automatic ultrasound report generation, leveraging a combination of unsupervised and supervised learning methods to aid the report generation process. Our framework incorporates unsupervised learning methods to extract potential knowledge from ultrasound text reports, serving as the prior information to guide the model in aligning visual and textual features, thereby addressing the challenge of feature discrepancy. Additionally, we design a global semantic comparison mechanism to enhance the performance of generating more comprehensive and accurate medical reports. To enable the implementation of ultrasound report generation, we constructed three large-scale ultrasound image-text datasets from different organs for training and validation purposes. Extensive evaluations with other state-of-the-art approaches exhibit its superior performance across all three datasets. Code and dataset are valuable at this link.

📄 PDF Abstract BibTeX arXiv:2406.00644

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

U2-BENCH: Benchmarking Large Vision-Language Models on Ultrasound Understanding

2025-05-23 · Anjie Le, Henan Liu, Yue Wang, Zhenyu Liu 외

Ultrasound is a widely-used imaging modality critical to global healthcare, yet its interpretation remains challenging due to its varying image quality on operators, noises, and anatomical structures. Although large visi…

BenchmarkingSpatial ReasoningText Generation

EchoVLM: Dynamic Mixture-of-Experts Vision-Language Model for Universal Ultrasound Intelligence

2025-09-18 · Chaoyin She, Ruifang Lu, Lida Chen, Wei Wang 외 arxiv

Ultrasound imaging has become the preferred imaging modality for early cancer screening due to its advantages of non-ionizing radiation, low cost, and real-time imaging capabilities. However, conventional ultrasound diag…

MMOTU: A Multi-Modality Ovarian Tumor Ultrasound Image Dataset for Unsupervised Cross-Domain Semantic Segmentation

2022-07-14 · Qi Zhao, Shuchang Lyu, Wenpei Bai, Linghan Cai 외

Ovarian cancer is one of the most harmful gynecological diseases. Detecting ovarian tumors in early stage with computer-aided techniques can efficiently decrease the mortality rate. With the improvement of medical treatm…

DecoderDomain AdaptationSegmentationSemantic Segmentation+1

Breast Ultrasound Report Generation using LangChain

2023-12-05 · Jaeyoung Huh, Hyun Jeong Park, Jong Chul Ye

Breast ultrasound (BUS) is a critical diagnostic tool in the field of breast imaging, aiding in the early detection and characterization of breast abnormalities. Interpreting breast ultrasound images commonly involves cr…

DiagnosticText Generation

Factored Attention and Embedding for Unstructured-view Topic-related Ultrasound Report Generation

2022-03-12 · Fuhai Chen, Rongrong Ji, Chengpeng Dai, Xuri Ge 외

Echocardiography is widely used to clinical practice for diagnosis and treatment, e.g., on the common congenital heart defects. The traditional manual manipulation is error-prone due to the staff shortage, excess workloa…

Decision MakingMedical Report Generation