arXiv Computer Vision

Evaluating the Effects of Inter-Observer and Model Variability on Radiological Peritoneal Cancer Index Assessment

arXiv Machine Learning
Jul 2

Foundation Models vs. Radiomics for Lung Computed Tomography: A Benchmark of Feature Extractors, Classification Heads, and Segmentation Choices

arXiv:2607. 01001v1 Announce Type: cross Abstract: Radiomics is the established approach for CT-based lung cancer phenotyping, yet comparisons with foundation models rarely isolate contributions of feature extractor, classification head, and segmentation choice, or test cross-cohort robustness.

By Nils Neukirch, Martin Maurer, Nils Strodthoff
arXiv Machine Learning
Jul 28

Trustworthy Medical Segmentation: Uncertainty-Aware U-Net Evaluation Under Clinical Image Degradation

arXiv:2607. 22727v1 Announce Type: cross Abstract: Medical image segmentation models often report high benchmark accuracy under ideal imaging conditions, yet their failures under clinical degradation can be quiet: sensor noise, patient motion, low- resolution acquisition, and contrast variability may all alter model behavior without producing an obvious warning.

By Pranav Kaliaperumal, Manisha Kaliaperumal
Hugging Face Trending Papers
Jul 9

Metrics or Mirage? An Audit of Evaluation Inconsistencies in Colonoscopy Polyp Segmentation Benchmarks

Progress in colonoscopy polyp segmentation is routinely reported through leaderboard comparisons on a small set of public benchmarks. We argue that this apparent progress is difficult to verify: a systematic audit of \textbf{27 papers} published between 2015 and 2026 reveals three structural problems in how the community evaluates models.

arXiv AI
3d ago

Extending TotalSegmentator: Predicting Patient and Acquisition Characteristics from CT and MR Images

arXiv:2608.29348v1 Announce Type: new Abstract: Background: Patient details and acquisition metadata are important for clinical decisions, image quality control, and automated research pipelines, but...

By Jakob Wasserthal, Joshy Cyriac, Michael Bach, Kimia Mozahheb Yousefi, Minh-Son To, M\'at\'e Sik, C\'edric H\'emon, Thomas Weikert, Martin Segeroth
arXiv AI
Aug 19

Comprehensive framework for evaluation of deep neural networks in detection and quantification of lymphoma from PET/CT images: clinical insights, pitfalls, and observer agreement analyses

This study presents a clinically relevant framework for evaluating deep neural networks that segment lymphoma lesions in PET/CT images, addressing gaps such as out‑of‑distribution testing and comparison with expert annotators. Using 611 multi‑institutional cases, the authors assess four networks (ResUNet, SegResNet, DynUNet, SwinUNETR) with lesion‑specific metrics, detection criteria, and metabolic‑characteristic‑based thresholds, finding that models perform best on large, intense lesions. The work also demonstrates that network errors mirror those of physicians, highlighting shared challenges with small, faint lesions.

By Shadab Ahamed, Yixi Xu, Sara Kurkowska, Claire Gowdy, Joo H. O, Ingrid Bloise, Don Wilson, Patrick Martineau, Fran\c{c}ois B\'enard, Fereshteh Yousefirizi, Rahul Dodhia, Juan M. Lavista, William B. Weeks, Carlos F. Uribe, Arman Rahmim
arXiv Computer Vision
4d ago

Report Supervision

The paper introduces Report Supervision (R‑Super), a framework that uses radiology reports to supervise tumor segmentation models. By incorporating loss functions that align segmentation outputs with report‑derived tumor counts, sizes, and locations, R‑Super improves detection and segmentation performance. Experiments on kidney and pancreatic tumors show up to a 15% increase in F1‑Score and DSC compared to mask‑only training, outperforming methods like CLIP and multi‑task learning.

By Pedro R. A. S. Bassia, Wenxuan Li, Jakob Wasserthal, Jieneng Chen, Xinze Zhou, Zheren Zhu, Chuntung Zhuanga, Sergio Decherchi, Andrea Cavalli, Kang Wang, Yang Yang, Alan Yuille, Zongwei Zhou
arXiv AI
Jun 9

Robust Renal Mass Segmentation on CT: A Validation Study of an AI-Based Framework

arXiv:2505. 07573v2 Announce Type: replace-cross Abstract: Renal mass segmentation has important potential to enhance the clinical workflow, especially in settings requiring quantitative assessments.

By Sarah de Boer, Hartmut H\"antze, Kiran Vaidhya Venkadesh, Myrthe A. D. Buser, Gabriel E. Humpire Mamani, Lina Xu, Lisa C. Adams, Jawed Nawabi, Keno K. Bressem, Bram van Ginneken, Mathias Prokop, Alessa Hering
arXiv AI
Aug 11

Performance of large language models in the optical diagnosis of colorectal polyps

arXiv:2608. 07543v1 Announce Type: cross Abstract: Background and Study Aims: Accurate optical diagnosis of colorectal polyps guides resection strategy and surveillance, with multimodal large language models (MLLMs) showing potential for image-based diagnosis.

By Joshua C. Vences, William T. Tran, Nikko Gimpaya, Catharine M. Walsh, Rishad J. Khan, Robert Bechara, Asher C. Wiggins, Celine N. Rousan, Kaitlyn V. G. L. Morgado, Angie Ibrahim, Kevin H. M. Kuo, Daniel von Renteln, Alexander Hann, Dennis L. Shung, Michael A. Scaffidi, Charles M\'enard, Joshua Landy, Samir C. Grover