paper-with-me

Papers

MMSummary: Multimodal Summary Generation for Fetal Ultrasound Video

2024-08-07 · Xiaoqing Guo, Qianhui Men, J. Alison Noble

We present the first automated multimodal summary generation system, MMSummary, for medical imaging video, particularly with a focus on fetal ultrasound analysis. Imitating the examination process performed by a human sonographer, MMSummary is designed as a three-stage pipeline, progressing from keyframe detection to keyframe captioning and finally anatomy segmentation and measurement. In the keyframe detection stage, an innovative automated workflow is proposed to progressively select a concise set of keyframes, preserving sufficient video information without redundancy. Subsequently, we adapt a large language model to generate meaningful captions for fetal ultrasound keyframes in the keyframe captioning stage. If a keyframe is captioned as fetal biometry, the segmentation and measurement stage estimates biometric parameters by segmenting the region of interest according to the textual prior. The MMSummary system provides comprehensive summaries for fetal ultrasound examinations and based on reported experiments is estimated to reduce scanning time by approximately 31.5%, thereby suggesting the potential to enhance clinical workflow efficiency.

📄 PDF Abstract BibTeX arXiv:2408.03761

Code (0)

등록된 구현이 없습니다.

Tasks

AnatomyLanguage ModelingLanguage ModellingLarge Language Model

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically
Focus 설명 없음

Similar Papers 제목 키워드 기반

FetalCLIP: A Visual-Language Foundation Model for Fetal Ultrasound Image Analysis

2025-02-20 · Fadillah Maani, Numan Saeed, Tausifa Saleem, Zaid Farooq 외

Foundation models are becoming increasingly effective in the medical domain, offering pre-trained models on large datasets that can be readily adapted for downstream tasks. Despite progress, fetal ultrasound images remai…

Age EstimationBenchmarking

A Review on Deep-Learning Algorithms for Fetal Ultrasound-Image Analysis

2022-01-28 · Maria Chiara Fiorentino, Francesca Pia Villani, Mariachiara Di Cosmo, Emanuele Frontoni 외

Deep-learning (DL) algorithms are becoming the standard for processing ultrasound (US) fetal images. Despite a large number of survey papers already present in this field, most of them are focusing on a broader area of m…

Deep LearningMedical Image Analysisparameter estimation

FetalFlex: Anatomy-Guided Diffusion Model for Flexible Control on Fetal Ultrasound Image Synthesis

2025-03-19 · Yaofei Duan, Tao Tan, Zhiyuan Zhu, Yuhao Huang 외

Fetal ultrasound (US) examinations require the acquisition of multiple planes, each providing unique diagnostic information to evaluate fetal development and screening for congenital anomalies. However, obtaining a compr…

AnatomyAnomaly DetectioncounterfactualDiagnostic+1

Towards Reliable Fetal Ultrasound Interpretation with Multi-Agent Collaboration

2026-05-25 · Xiaotian Hu, Mingxuan Liu, Junwei Huang, Kasidit Anmahapong 외 arxiv

Automated fetal ultrasound interpretation requires a workflow from visual perception, including plane recognition and anatomical segmentation, to clinical understanding, including biometric measurement and diagnostic rep…

Visual Question AnsweringVideo SummarizationImage Captioning

Epistemic-aware Vision-Language Foundation Model for Fetal Ultrasound Interpretation

2025-10-14 · Xiao He, Huangxuan Zhao, Guojia Wan, Wei Zhou 외 arxiv

Recent medical vision-language models have shown promise on tasks such as VQA, report generation, and anomaly detection. However, most are adapted to structured adult imaging and underperform in fetal ultrasound, which p…

Reinforcement LearningAnomaly Detection