paper-with-me

홈 › Papers

A multi-centre, multi-device benchmark dataset for landmark-based comprehensive fetal biometry

2025-12-18 · Chiara Di Vece, Zhehua Mao, Netanell Avisdris, Brian Dromey, Raffaele Napolitano, Dafna Ben Bashat, Francisco Vasconcelos, Danail Stoyanov, Leo Joskowicz, Sophia Bano arxiv

Accurate fetal growth assessment from ultrasound (US) relies on precise biometry measured by manually identifying anatomical landmarks in standard planes. Manual landmarking is time-consuming, operator-dependent, and sensitive to variability across scanners and sites, limiting the reproducibility of automated approaches. There is a need for multi-source annotated datasets to develop artificial intelligence-assisted fetal growth assessment methods. To address this bottleneck, we present an open, multi-centre, multi-device benchmark dataset of fetal US images with expert anatomical landmark annotations for clinically used fetal biometric measurements. These measurements include head bi-parietal and occipito-frontal diameters, abdominal transverse and antero-posterior diameters, and femoral length. The dataset comprises 4,513 de-identified US images from 1,904 subjects acquired at three clinical sites using seven different US devices. We provide standardised, subject-disjoint train/test splits, evaluation code, and baseline results to enable fair and reproducible comparison of methods. Using an automatic biometry model, we quantify domain shift and demonstrate that training and evaluation confined to a single centre substantially overestimate performance relative to multi-centre testing. To the best of our knowledge, this is the first publicly available multi-centre, multi-device, landmark-annotated dataset that covers all primary fetal biometry measures, providing a robust benchmark for domain adaptation and multi-centre generalisation in fetal biometry and enabling more reliable AI-assisted fetal growth assessment across centres. All data, annotations, training code, and evaluation pipelines are made publicly available.

📄 PDF Abstract BibTeX arXiv:2512.16710

Code (0)

등록된 구현이 없습니다.

Tasks

Domain Adaptation

Similar Papers 제목 키워드 기반

A comprehensive multimodal dataset and benchmark for ulcerative colitis scoring in endoscopy

2026-03-15 · Noha Ghatwary, Jiangbei Yue, Ahmed Elgendy, Hanna Nagdy 외 arxiv

Ulcerative colitis (UC) is a chronic mucosal inflammatory condition that places patients at increased risk of colorectal cancer. Colonoscopic surveillance remains the gold standard for assessing disease activity, and rep…

Image Captioning

Improving AI Efficiency in Data Centres by Power Dynamic Response

2025-10-13 · Andrea Marinoni, Sai Shivareddy, Pietro Lio', Weisi Lin 외 arxiv

The steady growth of artificial intelligence (AI) has accelerated in the recent years, facilitated by the development of sophisticated models such as large language models and foundation models. Ensuring robust and relia…

Interpretability-guided Data Augmentation for Robust Segmentation in Multi-centre Colonoscopy Data

2023-08-30 · Valentina Corbetta, Regina Beets-Tan, Wilson Silva

Multi-centre colonoscopy images from various medical centres exhibit distinct complicating factors and overlays that impact the image content, contingent on the specific acquisition centre. Existing Deep Segmentation net…

Data AugmentationImage SegmentationSegmentationSemantic Segmentation

Disease classification of macular Optical Coherence Tomography scans using deep learning software: validation on independent, multi-centre data

2019-07-11 · Kanwal K. Bhatia, Mark S. Graham, Louise Terry, Ashley Wood 외

Purpose: To evaluate Pegasus-OCT, a clinical decision support software for the identification of features of retinal disease from macula OCT scans, across heterogenous populations involving varying patient demographics, …

Benchmarking machine learning models on multi-centre eICU critical care dataset

2019-10-02 · Seyedmostafa Sheikhalishahi, Vevake Balaraman, Venet Osmani

Progress of machine learning in critical care has been difficult to track, in part due to absence of public benchmarks. Other fields of research (such as computer vision and natural language processing) have established …

BenchmarkingBIG-bench Machine LearningDecompensationMortality Prediction+1