Deep learning-based computer-aided diagnosis (CAD) systems have shown strong performance in breast cancer diagnosis, particularly for classification tasks in mammography. However, domain shifts across multi-site datasets remain a challenge, especially when models are applied to unseen domains.
arXiv:2607. 10358v1 Announce Type: cross Abstract: Foundation models are increasingly used as image feature extractors for mammography, but their robustness under external domain shift remains unclear.
By Giang Nguyen, Raghav Mehta, Emma A. M. Stanley, Tian Xia, Thi Hao Nguyen, Hieu Pham, Ben Glocker
arXiv:2512. 17605v2 Announce Type: replace-cross Abstract: Robust mammography registration is essential for clinically relevant applications like tracking disease progression in breast tissue.
By Svetlana Krasnova, Emiliya Starikova, Ilia Naletov, Andrey Krylov, Dmitry Sorokin
This study evaluates the classification accuracy of six modern deep‑learning architectures—VGG19, ResNet50, GoogleNet, ConvNeXt, EfficientNet, and Vision Transformers—on breast ultrasound images categorized by BI‑RADS. Using 2,945 training images and 936 validation images from 1,540 patients, the models were tested in full fine‑tuning, linear evaluation, and training‑from‑scratch settings. The best performance was achieved with full fine‑tuning, yielding 76.39 % accuracy and a 67.94 % F1 score.
By Malitha Gunawardhana, Norbert Zolek
arXiv:2607. 10188v1 Announce Type: cross Abstract: Breast cancer remains the most commonly diagnosed malignancy among women worldwide, yet accurate detection and characterization of breast masses in mammography remain challenging due to subtle intensity variations, heterogeneous tissue densities, and indistinct lesion boundaries that complicate radiological interpretation.
By Abu Fatema Mohammad Abdun Noor, Md Imam Ahasan, Md Samiul Ahasan, Kah Ong Michael Goh, S M Hasan Mahmud, Raihana Zannat
arXiv:2607. 11343v1 Announce Type: cross Abstract: Accurate breast cancer risk prediction from screening mammography is critical for enabling personalized screening intervals and early detection.
By Solveig Thrun, Zijun Sun, Suaiba A. Salahuddin, Kristoffer Wickstr{\o}m, Elisabeth Wetzer, Stine Hansen, Robert Jenssen, Michael Kampffmeyer
Accurate breast cancer classification from mammography requires effective integration of complementary information from craniocaudal (CC) and mediolateral oblique (MLO) views, which provide a more complete characterization of breast abnormalities. However, existing multi-view learning approaches typically rely on feature-level aggregation or single-stage cross-attention, which can entangle view-specific and shared representations and restrict interaction to limited network depths.
arXiv:2607. 06309v1 Announce Type: cross Abstract: Accurate breast cancer classification from mammography requires effective integration of complementary information from craniocaudal (CC) and mediolateral oblique (MLO) views, which provide a more complete characterization of breast abnormalities.
By Aysan Ghayouri Pirsoltan, Shima Babakordi, Mohammad Reza Mohammadi
MagViT is an interpretable multi‑magnification transformer that classifies breast histopathology images by extracting representations from four BreakHis magnifications (40X, 100X, 200X, 400X) and fusing them with a learnable, scale‑gated mechanism that can mask missing scales. The model selects the most accurate architectural branch at the patient level using five‑fold cross‑validation, achieving high performance on BreakHis (mean image accuracy 0.9191, patient accuracy 0.9643, macro‑F1 0.9042) and demonstrating preliminary cross‑dataset generalization on BUSI and IDC. Grad‑CAM visualizations confirm that the network focuses on diagnostically relevant regions across magnifications.
By Nabil Ashab, Soumit Kumar Kundu, Saif Mahmud Parvez, Shahadat Hossain Sohag, Bidhan Biswas, Nazmus Subha
arXiv:2608. 15019v1 Announce Type: cross Abstract: Breast mass segmentation is an important step in computer-aided mammography, but it remains difficult because masses can have low contrast, irregular shapes, and boundaries that blend with surrounding breast tissue.
By Alibek Kamiluly, Milana Muratova, Yash Patel, Fan Li
M3D‑Net is a mammography encoder that hierarchically coordinates multi‑scale coordinate attention, bounded dynamic feature reuse, and differential attention through resolution‑aware operator placement. It preserves earlier features within stages, integrates local and global context via coordinate‑aware aggregation, and applies differential attention at coarse resolutions. In image‑only classification on AISSLab mammography and an adapted image‑clinical model on BrEaST ultrasound, M3D‑Net achieves the highest validation accuracy and lowest endpoint cross‑entropy loss compared to EdgeNeXt, RepViT, and TransXNet, with accuracies of 97.78% and 80.39% respectively.
By Zheng Yu, Xinhang Li, Jiabao Gao, Boyang Wang, Xiang Li
The paper introduces TopKSigLIP, a vision‑language model tailored for mammography that tackles two key challenges: high‑resolution imaging and homogeneous radiology reports. It replaces standard CLIP training with a TopK‑Patch module that selects sparse high‑resolution patches likely to contain lesions, and a Sup‑sigmoid loss that uses soft labels from structured data instead of contrastive loss. TopKSigLIP outperforms existing open‑source mammography and general medical VLMs on zero‑shot tasks such as density assessment, BI‑RADS classification, finding subtyping, and cancer prediction, while also providing better lesion localization than Grad‑CAM.
By Young Seok Jeon, Beatrice Brown-Mulry, Rohan Satya Isaac, Anjana Dissanayaka, Theo Dapamede, Mohammadreza Chavoshi, Judy Gichoya, Hari Trivedi