ProtoCAM is an explainable few‑shot learning framework for classifying breast lesions in ultrasound images. It combines mask‑guided feature encoding, prototypical metric learning, and gradient‑based visual explanations to leverage limited annotated data. Evaluated on the BUSI dataset, ProtoCAM achieved a macro F1‑score of 0.910 in a 3‑way 5‑shot setting, outperforming standard supervised CNNs, with ResNet18 reaching 91.65% under 15‑shot conditions.
By Ashkan Ebadi
arXiv:2608.24281v1 Announce Type: new
Abstract: Reducing annotation requirements remains a key challenge in developing robust medical object detectors. To address this, Vision-Language (VL) object de...
By Sheethal Bhat, Bogdan Georgescu, Awais Mansoor, Mathias Zinnen, Pranjal Sahu, Florin C. Ghesu, Sasa Grbic, Andreas Maier
The CXR‑LT 2026 Challenge introduces a multi‑center, long‑tailed chest X‑ray classification benchmark with over 145,000 radiologist‑annotated images from PadChest and NIH datasets. It defines two core tasks: robust multi‑label classification on 30 known classes and open‑world generalization to 6 unseen rare disease classes. The paper outlines data collection, annotation, solution strategies, and evaluates performance across head‑vs‑tail, calibration, and cross‑center gaps, noting that vision‑language models improve in‑distribution and zero‑shot performance but rare‑finding detection under multi‑center shift remains difficult.
By Hexin Dong, Yi Lin, Pengyu Zhou, Fengnian Zhao, Alan Clint Legasto, Juno Cho, Dohui Kim, Justin Namuk Kim, Mingeon Kim, Sunwoo Kwak, Gabriel Moy\`a-Alcover, Ky Trung Nguyen, Thanh-Huy Nguyen, Ha-Hieu Pham, Huy-Hieu Pham, Huy Le Pham, Nikhileswara Rao Sulake, Aina Tur-Serrano, Ruichi Zhang, Ang Zu, Adam E. Flanders, Zhiyong Lu, Ronald M. Summers, Mingquan Lin, Hao Chen, Yuzhe Yang, George Shih, Yifan Peng
arXiv:2607. 20641v1 Announce Type: new Abstract: Federated learning (FL) enables multiple clinical institutions to collaboratively train a shared disease classifier without centralizing patient data.
By Afsaneh Mahanipour, Hana Khamfroush
The paper introduces MedIDL, a Medical Imaging Disentanglement Learning framework that separates disease-related features from confounding covariates and individual variability in medical images. It achieves this by projecting image features into three orthogonal latent spaces—disease classification, covariate alignment, and a Gaussian head for individual variation—using specialized disentanglement heads. Across seven diverse imaging datasets, MedIDL surpasses state‑of‑the‑art supervised and self‑supervised methods in classification accuracy, and its latent representations and gradient‑based visualizations align with known clinical patterns.
By Shengjie Zhang, Jinglin Zhang, Zhuangzhuang Jiang, Ziqi Yu, Yipin Zhang, Qi Zhang, Xiang Chen, Haibo Yang, Fei Gao, Longbiao Cui, Yuan Zhou, Xiao-Yong Zhang, Alzheimer's Disease Neuroimaging Initiative
arXiv:2607. 00975v1 Announce Type: cross Abstract: Chest X-ray multi-label classification is a core task in intelligent medical imaging diagnosis.
By Tong Shao, Hongshun Ling, Li Zhang, Jinjing Wu, Junke Wang, Yuan Gao, Fang Wang