arXiv AI

FADA: Accessible fetal ultrasound interpretation and annotation with a selectively distilled unified vision-language model

arXiv:2606. 11106v1 Announce Type: cross Abstract: A global shortage of trained sonographers limits prenatal ultrasound screening in low- and middle-income countries, where over half of pregnant women receive no skilled sonography.

arXiv Computer Vision
Sep 18

Open ultrasound foundation model for robust segmentation and clinical measurement across heterogeneous settings

The paper introduces SonoCorpus, an open dataset of 456,963 ultrasound images with 1,626,085 expert masks from 53 public sources across 24 clinical applications and 17 countries, and SonoBase, an interactive segmentation foundation model pretrained on this data. SonoBase outperforms existing models (SAM2, MedSAM2, MedSAM3) on fifteen diverse evaluation datasets, matching specialist models and achieving clinically relevant accuracy for metrics such as ejection fraction, fetal head circumference, and gestational age. The authors provide full reproducibility resources, including checkpoints, optimizer states, and starter code, to enable community adoption and further development.

By Chao Qin, Fahad Shahbaz Khan, Salman Khan, Sarim Ather, Siddiq Anwar, Rao Muhammad Anwer, Shadab Khan
arXiv AI
Jul 22

FedCC: A Low-Resource Federated Adaptation of Foundation Models for Robust Corpus Callosum localization in Fetal Ultrasound Images

arXiv:2607. 18283v1 Announce Type: cross Abstract: Accurate localization of the corpus callosum (CC) in fetal ultrasound (US) images is crucial for the early identification of neurodevelopmental abnormalities.

By Alessandro Di Matteo, Sara Moccia, Giuseppe Rizzo, Gianpaolo Grisolia, Ricciarda Raffaelli, Lorenzo Vasciaveo, Francesco D'Antonio, Maria Chiara Fiorentino
arXiv Computer Vision
Sep 15

Deep Learning-based Intelligent Diagnosis of Congenital Uterine Anomalies in 3D Ultrasound

arXiv:2609.15225v1 Announce Type: new Abstract: Objective: To develop an intelligent framework, termed CUA-Net, for the automated classification of congenital uterine anomalies (CUA) without requirin...

By Yueyue Xu, Yuhao Huang, Jiaxiao Deng, Yuanji Zhang, Haoming Zhang, Jiajia Qu, Shiying Zheng, Xiaomei Tang, Haining Chen, Chengcai Chen, Yiyi Wu, Xin Yang, Dong Ni
arXiv AI
Jun 16

Enabling Real-Time Point-of-Care Ultrasound Segmentation: A GPU-Free Deployment in Resource-Limited Settings

arXiv:2606. 15176v1 Announce Type: cross Abstract: Ultrasound imaging is the most widely adopted medical modality globally due to its low cost and portability, yet artificial intelligence (AI) deployment remains constrained by reliance on GPU-accelerated models, creating a structural paradox where the cost of "intelligence" exceeds that of the imaging device itself.

By Weihao Gao
arXiv Computer Vision
Sep 18

FreqDINO++: A Frequency-Guided Multi-Task Routing Vision Foundation Model for Universal Ultrasound Analysis

FreqDINO++ is a frequency‑guided multi‑task routing vision foundation model designed for universal ultrasound analysis. It introduces a Multi‑task Routing Adapter for efficient task‑common and task‑specific integration, a Frequency‑aware Feature Enhancer to capture multi‑scale frequency characteristics, and a Task‑aligned Collaborative Decoder that promotes collaboration between dense and global prediction tasks. Experiments on large‑scale multi‑task and external single‑task ultrasound benchmarks show that FreqDINO++ outperforms strong baselines and recent foundation models across 27 diverse clinical task scenarios, with promising generalization to unseen data.

By Qing Xu, Yixuan Zhang, Yue Li, Xiangjian He, Qian Zhang, Mainul Haque, Rong Qu, Wenting Duan, Jieyun Bai, Zhen Chen
arXiv Computer Vision
Aug 28

Anatomy-Guided Foundation Model Adaptation with Within-Case Prototype Supervision for Standard Plane Detection in Fetal Ultrasound Blind Sweeps

AnatoProto is a lightweight sequence‑level framework that adapts a frozen BiomedCLIP encoder for detecting the fetal abdominal circumference standard plane in low‑cost obstetric blind sweeps. It incorporates anatomy‑weighted spatial pooling, a within‑case prototype loss, a three‑stage cascade refinement, and a hybrid prediction head to address the highly imbalanced, short‑segment nature of the task. On the ACOUSLIC‑AI benchmark, AnatoProto achieves a test F1 of 67.72, surpassing the best foundation‑model baseline by 13.20 F1 and the best video temporal‑action‑detection baseline by 15.76 F1.

By Yuzhe Zhao
arXiv Computer Vision
Aug 27

Less Contouring, More Accuracy: Lesion-Guided ROI Deep Learning for Ovarian Ultrasound Classification

The study evaluates lesion‑guided region‑of‑interest (ROI) deep learning for ovarian ultrasound classification, comparing it to global image, lesion contour, and contour‑based radiomics approaches across two public datasets. Using four deep‑learning architectures, the lesion‑guided ROI strategy achieved the highest accuracy (93.10% on MMOTU and 97.56% on OUD) with an AUC of 0.99, while requiring less annotation effort than contour‑based methods.

By Mehran Ahmad, Ali Abbasian Ardakani, Afshin Mohammadi, Alisa Mohebbi, Gernot Kronreif, Sepideh Hatamikia
arXiv AI
Aug 25

Robust Lightweight Deep Learning Models for Oral Cancer Screening

arXiv:2608.21583v1 Announce Type: new Abstract: Oral cancer is a leading cause of mortality in low-to-middle-income countries, where a shortage of specialists delays diagnosis. While point-of-care sc...

By Siddhant Bharadwaj, Aakash Shedsale, Tejashree Subramanya, Mohd. Azfar, Praveen Birur, Debnath Pal, Shankararama Sharma, Anupama Shetty, Rajesh Sundaresan