arXiv Machine Learning

Toward Efficient Weakly Supervised Semantic Segmentation Using Only Low-Magnification Histopathological Images

arXiv:2607. 10783v1 Announce Type: cross Abstract: Whole-slide images (WSIs) provide rich tissue-level and cellular-level information, but storing and transmitting high-magnification pathology data is resource-intensive.

arXiv Computer Vision
Sep 3

AtlasPatch: Scalable Foundation Model-based Tissue Detection and Patch Extraction for Computational Pathology

AtlasPatch is a scalable, high‑throughput whole‑slide image preprocessing method that uses a foundation‑model‑based tissue detector operating at thumbnail resolution. By updating only 0.076% of the SAM2 model weights and leveraging a curated dataset of 30,000 thumbnail‑mask pairs, it generates accurate tissue masks and directly produces patch coordinates at the desired magnification, eliminating repeated patch‑level inference. The approach achieves 0.986 precision, is up to 16× faster than existing deep‑learning methods, and maintains downstream multiple‑instance learning performance across six slide‑level classification tasks.

By Ahmed Alagha, Christopher Leclerc, Yousef Kotp, Omar Metwally, Calvin Moras, Peter Rentopoulos, Ghodsiyeh Rostami, Bich Ngoc Nguyen, Jumanah Baig, Abdelhakim Khellaf, Vincent Quoc-Huy Trinh, Rabeb Mizouni, Hadi Otrok, Jamal Bentahar, Mahdi S. Hosseini
arXiv Computer Vision
Sep 22

Patch-to-Global: Random Patch Diffusion for Globally Consistent Megapixel Artifact Inpainting in Whole Slide Images

arXiv:2609.24116v1 Announce Type: new Abstract: Although deep learning has advanced Whole Slide Image (WSI) Analysis, tissue artifacts like bubbles and folds often cause silent failures by concealing...

By Hyeseong Lee, Eunsu Kim, D M Bappy, Ho Heon Kim, Youngsuk Lee, Se Young Chun, Jang-Hwan Choi, Sung Hak Lee, Sangjeong Ahn
arXiv AI
Aug 28

Pixel Wised Lesion Prediction on COVID-19 CT Imagery: A Comparative Analysis of Automated Image Segmentation Architectures

The study evaluates four deep‑learning segmentation architectures—Unet, PSPNet, Linknet, and FPN—paired with six pre‑trained encoders to predict COVID‑19 lesions in CT images. Experiments on three COVID‑19 CT datasets show high accuracy, achieving a maximum binary F1‑score of 98% and multi‑class F1‑scores of 75% and 77%. The work aims to provide a standardized performance benchmark for medical image segmentation and a reference for other imaging scenarios.

By Sarmad Khan, Basim Azam, Arslan Shaukat
arXiv AI
Jun 17

Enhancing Pathological VLMs with Cross-scale Reasoning

arXiv:2606. 17412v1 Announce Type: cross Abstract: Pathological images are inherently multi-scale, requiring pathologists to integrate evidence from global tissue architecture at low magnification to cellular morphology at higher magnification for accurate diagnosis.

By Chi Phan, Tianyi Zhang, Qiaochu Xue, Yufeng Wu, Dan Hu, Zeyu Liu, Sudong Wang, Yueming Jin