arXiv Computer Vision

SegKAN: High-Resolution Medical Image Segmentation with Long-Distance Dependencies

SegKAN is a new model for high‑resolution medical image segmentation that tackles fragmentation and noise in hepatic vessel CT scans. It replaces the standard embedding module with a novel convolutional network to smooth noise and avoid gradient explosion, and reinterprets spatial relationships between Patch blocks as temporal relationships to better capture positional dependencies. Experiments on a hepatic vessel dataset show a 1.78% improvement in Dice score over the current state‑of‑the‑art model, indicating that the new structure enhances segmentation performance for extended objects.

arXiv Computer Vision
Aug 27

Improving Cross-Site Whole-Heart Segmentation

The paper presents a modality‑routed 3D cardiac segmentation pipeline that combines TotalSegmentator‑initialized nnU‑Netv2 models with site‑characterized, label‑preserving appearance augmentation. By analyzing measurable image properties across sites, the authors design a bias‑field plus Bezier augmentation strategy that smooths spatial intensity perturbations and remaps intensities nonlinearly, followed by class‑wise largest‑connected‑component cleanup. On held‑out validation splits, this approach raises CT mean Dice from 0.8350 to 0.9135 and MRI mean Dice from 0.7695 to 0.7830 while reducing HD95, demonstrating improved cross‑site robustness in limited‑data whole‑heart segmentation.

By Tanish Mudaliar, Justin Li, Daniel Lin, Julianna Vo, Kaitao Liao, Xin Wang, Shu Hu
arXiv AI
Jul 31

PatchDenoiser: Parameter-efficient multi-scale patch learning and fusion denoiser for Low-dose CT imaging

arXiv:2602. 21987v3 Announce Type: replace-cross Abstract: Low-dose CT images are essential for reducing radiation exposure in cancer screening, pediatric imaging, and longitudinal monitoring protocols, but their quality is often degraded by noise from low-dose acquisition, patient motion, or scanner limitations, affecting both clinical interpretation and downstream analysis.

By Jitindra Fartiyal, Pedro Freire, Sergei K. Turitsyn, Sergei G. Solovski
arXiv AI
6d ago

ReG-SAM: Reference Graph-Driven SAM for 2D Foundational Vessel Segmentation

ReG-SAM is a SAM-based framework designed for 2D vessel segmentation in medical images. It introduces reference graph prompt embeddings (GPEs) and vascular prototype embeddings (VPEs) to capture global spatial and fine-grained modality-specific vessel features, respectively. By building a modality-wise vascular database and learning these embeddings from reference masks, ReG-SAM consistently outperforms existing baselines across 19 datasets, especially on thin vessels.

By Donghang Lyu, Zichen Zhang, Oleh Dzyubachyk, Marius Staring
arXiv Machine Learning
Sep 14

Bridging Vision Foundation Model Priors with CLIP for Spatial-aware Few-shot Anomaly Detection in Medical Images

The paper introduces Spatial‑FAD, a few‑shot medical anomaly detection framework that fuses Vision‑Language Model (CLIP) semantics with spatial priors from Vision Foundation Models (DINO). A VFM‑enhanced adapter injects structural affinity into CLIP features, while a sliding‑window aggregation produces high‑resolution embeddings for finer lesion localization. Prototype‑enhanced support memory further improves efficiency and performance, yielding significant gains on Liver CT, Retinal OCT, and Brain MRI datasets, notably an 11.4% Dice improvement in 4‑shot scenarios.

By Juzheng Miao, Yuchen Yuan, Cheng Chen, Pheng-Ann Heng