arXiv AI

Externally Validated Breast Ultrasound Segmentation via Multi-task Learning with BI-RADS-Consistent Morphological Priors

arXiv:2511. 15968v2 Announce Type: replace-cross Abstract: External validation of breast ultrasound segmentation models remains limited because internal train--test splits do not capture domain shifts across imaging systems, acquisition protocols, and patient populations.

arXiv Computer Vision
Sep 18

Performance of Machine Learning Classification in Sonomammogram Images using BI-RADS

This study evaluates the classification accuracy of six modern deep‑learning architectures—VGG19, ResNet50, GoogleNet, ConvNeXt, EfficientNet, and Vision Transformers—on breast ultrasound images categorized by BI‑RADS. Using 2,945 training images and 936 validation images from 1,540 patients, the models were tested in full fine‑tuning, linear evaluation, and training‑from‑scratch settings. The best performance was achieved with full fine‑tuning, yielding 76.39 % accuracy and a 67.94 % F1 score.

By Malitha Gunawardhana, Norbert Zolek
arXiv Machine Learning
Jul 14

BiLoG-Net: A Bi-Context Location-Guided Network for Breast Mass Segmentation and Malignancy Classification in Mammography

arXiv:2607. 10188v1 Announce Type: cross Abstract: Breast cancer remains the most commonly diagnosed malignancy among women worldwide, yet accurate detection and characterization of breast masses in mammography remain challenging due to subtle intensity variations, heterogeneous tissue densities, and indistinct lesion boundaries that complicate radiological interpretation.

By Abu Fatema Mohammad Abdun Noor, Md Imam Ahasan, Md Samiul Ahasan, Kah Ong Michael Goh, S M Hasan Mahmud, Raihana Zannat
arXiv AI
Aug 25

SAS: Segment Anything Small for Ultrasound -- A Non-Generative Data Augmentation Technique for Robust Deep Learning in Ultrasound Imaging

The paper introduces Segment Anything Small (SAS), a data‑augmentation method that improves deep‑learning segmentation of small anatomical structures in ultrasound images. SAS uses two transformations: resizing and embedding organ thumbnails into a black background to vary organ scale, and adding noise to regions of interest to mimic tissue texture variability. Experiments on one internal and five external datasets show Dice score gains up to 0.35, with an average improvement of 0.16, and demonstrate that SAS enhances model robustness and generalizability without adding hallucinations or artifacts.

By Danielle L. Ferreira, Ahana Gangopadhyay, Hsi-Ming Chang, Ravi Soni, Gopal Avinash
arXiv Computer Vision
Sep 4

Improving Clinical Target Volume Segmentation Accuracy using Anatomical Priors and Active Learning for the AGITG TOPGEAR Clinical Trial

The study explores how adding anatomical priors and active learning can improve the accuracy of deep learning models for segmenting the Clinical Target Volume (CTV) in gastric cancer radiotherapy. Using 100 retrospective CT scans, an nnU‑Net model trained on 10 expert‑contoured cases was enhanced with voxel‑wise anatomical prior maps and iterative active learning over four rounds. The combined approach raised the mean Dice Similarity Coefficient from 0.84 to 0.87, demonstrating that both techniques individually and together improve segmentation performance and generalizability.

By Phillip Chlap, Mark Lee, Trevor Leong, Matthew Field, Jason Dowling, Hang Min, Julie Chu, Jennifer Tan, Phillip K. Tran, Tomas Kron, Annette Haworth, Martin A. Ebert, Shalini K. Vinod, Lois Holloway
Hugging Face Trending Papers
Jul 21

MIRAGE: Multi-scale Lesion-Informed Representation with Auxiliary Guidance for MRI Contrast Enhancement

Inferring contrast enhancement from one pre-contrast breast MRI slice is underdetermined: post-contrast appearance contains physiological information that is not uniquely encoded in baseline anatomy. Optimizing only paired pixel fidelity can suppress uncertain lesion enhancement, whereas adversarial or stochastic generative objectives can favor realistic post-contrast appearance without guaranteeing patient-specific lesion fidelity.

arXiv AI
Sep 1

Subtraction-Based Tumor Segmentation and Lesion-Centered pCR Prediction for the MAMA-MIA Challenge

Team FME submitted a method for the MAMA-MIA Challenge that tackles primary tumor segmentation and pathological complete response (pCR) prediction using dynamic contrast‑enhanced breast MRI. For segmentation, they employed a five‑fold residual‑encoder nnU‑Net ensemble trained on the first post‑contrast minus pre‑contrast image, augmented with mirroring test‑time augmentation and largest‑connected‑component filtering, achieving a Dice score of 0.713 and a normalized Hausdorff distance of 0.099. For pCR prediction, they ensembled 25 pretrained 3D video classifiers on lesion‑centred crops from the pre‑contrast and first two post‑contrast volumes, reaching a balanced accuracy of 0.541 and an equalized‑odds disparity of 0.212, and ranked second in both tasks.

By Kai Geissler, Raphael Sch\"afer
arXiv Computer Vision
Aug 27

Less Contouring, More Accuracy: Lesion-Guided ROI Deep Learning for Ovarian Ultrasound Classification

The study evaluates lesion‑guided region‑of‑interest (ROI) deep learning for ovarian ultrasound classification, comparing it to global image, lesion contour, and contour‑based radiomics approaches across two public datasets. Using four deep‑learning architectures, the lesion‑guided ROI strategy achieved the highest accuracy (93.10% on MMOTU and 97.56% on OUD) with an AUC of 0.99, while requiring less annotation effort than contour‑based methods.

By Mehran Ahmad, Ali Abbasian Ardakani, Afshin Mohammadi, Alisa Mohebbi, Gernot Kronreif, Sepideh Hatamikia
arXiv Computer Vision
Sep 25

Integrating Local Detail and Global Context: A Dual-Input Multi-Task Learning Framework for Bone Tumor Diagnosis

The paper introduces a dual‑input, multi‑task learning framework that jointly segments and classifies bone tumors by applying bidirectional cross‑modal attention between a lesion crop and the full radiograph. Using a YOLO‑based detector and a dual‑stream DenseNet121 architecture, the model fuses fine‑grained lesion detail with global anatomical context through a novel cross‑modal attention fusion strategy and hierarchical multi‑scale feature fusion. On the multi‑institutional Bone Tumor X‑ray Radiograph Dataset, the approach outperforms single‑input baselines, achieving a Dice coefficient of 0.896 and a macro‑averaged F1‑score of 0.928, with an AUC of 0.999 for malignant osteosarcoma.

By S. M. Nasif Uddin, Rusab Sarmun, Muhammad E. H. Chowdhury, Adam Mushtak, Israa Al-Hashimi, Sohaib Bassam Zoghoul
arXiv Machine Learning
Aug 27

Unsupervised Anatomical Feature Learning via Diffusion Models: Enhanced Medical Image Segmentation with Denoising Diffusion Probabilistic Models

The paper introduces an unsupervised approach to medical image segmentation by training a Denoising Diffusion Probabilistic Model (DDPM) on 21 unlabeled abdominal CT scans to learn anatomical features. The encoder weights from the DDPM are transferred to a U‑Net for downstream segmentation on the BTCV multi‑organ dataset, resulting in a significant Dice score improvement for liver segmentation from 0.75 to 0.93. In low‑data regimes, diffusion‑pretrained models retain robust performance, achieving high Dice scores even with only 10% of labeled data.

By Akshat G, Divyansh Gupta, Shaleen Bhatnagar, Shilpa Ankalaki, Tusar Kanti Mishra
arXiv AI
Aug 17

CMCNet: Aligning Ultrasound Image Embeddings with Textual TI-RADS Representations for Fine-Grained Thyroid Classification

arXiv:2608. 13939v1 Announce Type: cross Abstract: Ultrasound is the primary imaging modality for assessing thyroid nodules, and the ACR TI-RADS framework standardizes diagnosis through five ultrasound feature categories that are aggregated into five risk levels (TR1-TR5).

By Bingxin Yu, Xueli Wang, Jerry Zhou, Wenyan Wang, Li Wen, Lan Huang, Xin Feng, Fengfeng Zhou, Kewei Li