A multi-centre, multi-device benchmark dataset for landmark-based comprehensive fetal biometry

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Di Vece, Chiara, Mao, Zhehua, Avisdris, Netanell, Dromey, Brian, Napolitano, Raffaele, Bashat, Dafna Ben, Vasconcelos, Francisco, Stoyanov, Danail, Joskowicz, Leo, Bano, Sophia
Formato: Preprint
Publicado: 2025
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866914595758669824
author Di Vece, Chiara
Mao, Zhehua
Avisdris, Netanell
Dromey, Brian
Napolitano, Raffaele
Bashat, Dafna Ben
Vasconcelos, Francisco
Stoyanov, Danail
Joskowicz, Leo
Bano, Sophia
author_facet Di Vece, Chiara
Mao, Zhehua
Avisdris, Netanell
Dromey, Brian
Napolitano, Raffaele
Bashat, Dafna Ben
Vasconcelos, Francisco
Stoyanov, Danail
Joskowicz, Leo
Bano, Sophia
contents Accurate fetal growth assessment from ultrasound (US) relies on precise biometry measured by manually identifying anatomical landmarks in standard planes. Manual landmarking is time-consuming, operator-dependent, and sensitive to variability across scanners and sites, limiting the reproducibility of automated approaches. There is a need for multi-source annotated datasets to develop artificial intelligence-assisted fetal growth assessment methods. To address this bottleneck, we present an open, multi-centre, multi-device benchmark dataset of fetal US images with expert anatomical landmark annotations for clinically used fetal biometric measurements. These measurements include head bi-parietal and occipito-frontal diameters, abdominal transverse and antero-posterior diameters, and femoral length. The dataset comprises 4,513 de-identified US images from 1,904 subjects acquired at three clinical sites using seven different US devices. We provide standardised, subject-disjoint train/test splits, evaluation code, and baseline results to enable fair and reproducible comparison of methods. Using an automatic biometry model, we quantify domain shift and demonstrate that training and evaluation confined to a single centre substantially overestimate performance relative to multi-centre testing. To the best of our knowledge, this is the first publicly available multi-centre, multi-device, landmark-annotated dataset that covers all primary fetal biometry measures, providing a robust benchmark for domain adaptation and multi-centre generalisation in fetal biometry and enabling more reliable AI-assisted fetal growth assessment across centres. All data, annotations, training code, and evaluation pipelines are made publicly available.
format Preprint
id arxiv_https___arxiv_org_abs_2512_16710
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle A multi-centre, multi-device benchmark dataset for landmark-based comprehensive fetal biometry
Di Vece, Chiara
Mao, Zhehua
Avisdris, Netanell
Dromey, Brian
Napolitano, Raffaele
Bashat, Dafna Ben
Vasconcelos, Francisco
Stoyanov, Danail
Joskowicz, Leo
Bano, Sophia
Computer Vision and Pattern Recognition
68-04, 68-11
D.0; E.0; I.2; I.4; I.5; J.2
Accurate fetal growth assessment from ultrasound (US) relies on precise biometry measured by manually identifying anatomical landmarks in standard planes. Manual landmarking is time-consuming, operator-dependent, and sensitive to variability across scanners and sites, limiting the reproducibility of automated approaches. There is a need for multi-source annotated datasets to develop artificial intelligence-assisted fetal growth assessment methods. To address this bottleneck, we present an open, multi-centre, multi-device benchmark dataset of fetal US images with expert anatomical landmark annotations for clinically used fetal biometric measurements. These measurements include head bi-parietal and occipito-frontal diameters, abdominal transverse and antero-posterior diameters, and femoral length. The dataset comprises 4,513 de-identified US images from 1,904 subjects acquired at three clinical sites using seven different US devices. We provide standardised, subject-disjoint train/test splits, evaluation code, and baseline results to enable fair and reproducible comparison of methods. Using an automatic biometry model, we quantify domain shift and demonstrate that training and evaluation confined to a single centre substantially overestimate performance relative to multi-centre testing. To the best of our knowledge, this is the first publicly available multi-centre, multi-device, landmark-annotated dataset that covers all primary fetal biometry measures, providing a robust benchmark for domain adaptation and multi-centre generalisation in fetal biometry and enabling more reliable AI-assisted fetal growth assessment across centres. All data, annotations, training code, and evaluation pipelines are made publicly available.
title A multi-centre, multi-device benchmark dataset for landmark-based comprehensive fetal biometry
topic Computer Vision and Pattern Recognition
68-04, 68-11
D.0; E.0; I.2; I.4; I.5; J.2
url https://arxiv.org/abs/2512.16710