arXiv Computer Vision

From Pixel Generation to Topological Inference: Structural Dual Super-Resolution for Trustworthy Cross-Physical-Domain Trabecular Morphology Learning

The paper introduces Structural Dual Super‑Resolution (SDN), a novel approach that shifts from pixel‑level super‑resolution to topological inference for trabecular bone morphology. By training on 2‑D slices and evaluating on 3‑D morphological metrics, SDN learns to predict invariant microstructures from low‑resolution CT inputs, using bidirectional modeling, a multi‑scale consistency discriminator, and four structural duality constraints. The method achieves SSIM of 0.8 and morphological parameters closely matching synchrotron micro‑CT across six metrics, demonstrating cross‑source generalization and trustworthy inference rather than mere pixel generation.

arXiv Computer Vision
4d ago

Cross-Modality Structural Guidance in 3D Latent Diffusion for Robust FLAIR Super-Resolution

The paper introduces MR‑DiffuSR, a 3‑D latent diffusion framework that uses high‑resolution T1w structural priors to guide super‑resolution of thick‑slice FLAIR MRI scans. By applying cross‑modality structural swin attention and a mixed‑scale degradation strategy, the method avoids hallucinations and remains robust across varying slice thicknesses. On ADNI datasets, MR‑DiffuSR outperforms CNN and 2‑D diffusion baselines, achieving high PSNR, SSIM, and low LPIPS, and maintains strong white‑matter hyperintensity segmentation performance even at 7 mm equivalent slice thickness.

By Haoyu Lan, Jiazhen Zhang, John Onofrey, Bino Varghese, Nasim Sheikh-Bahaei, Arthur W. Toga, Jeiran Choupan
arXiv AI
Jun 10

Deep Slice Interpolation for Reducing Through-Plane Anisotropy and Noise in Head CT

arXiv:2606. 09953v1 Announce Type: cross Abstract: Head computed tomography (CT) typically uses sub-millimeter in-plane resolution but 2-5 mm through-plane spacing, creating substantial anisotropy that degrades multiplanar reconstructions, volumetric measurements such as hematoma volume estimation, and downstream algorithms that assume near-isotropic voxels.

By Luis Cort\'es Ferre, Miguel A. Guti\'errez-Naranjo, Marcin Balcerzyk
arXiv AI
Aug 25

SAS: Segment Anything Small for Ultrasound -- A Non-Generative Data Augmentation Technique for Robust Deep Learning in Ultrasound Imaging

The paper introduces Segment Anything Small (SAS), a data‑augmentation method that improves deep‑learning segmentation of small anatomical structures in ultrasound images. SAS uses two transformations: resizing and embedding organ thumbnails into a black background to vary organ scale, and adding noise to regions of interest to mimic tissue texture variability. Experiments on one internal and five external datasets show Dice score gains up to 0.35, with an average improvement of 0.16, and demonstrate that SAS enhances model robustness and generalizability without adding hallucinations or artifacts.

By Danielle L. Ferreira, Ahana Gangopadhyay, Hsi-Ming Chang, Ravi Soni, Gopal Avinash
arXiv Computer Vision
Aug 31

Physics-Guided Flow Matching for CT Image Reconstruction

The paper introduces a high‑resolution Rectified Flow Matching model trained on 256×256 chest CT images to serve as a generative prior for CT reconstruction. A two‑stage training strategy—initial strong anatomically informed augmentation followed by fine‑tuning—helps mitigate overfitting and improve structural fidelity. When evaluated on various CT inverse problems, Flow Matching‑based reconstruction methods outperform diffusion‑based algorithms in PSNR, SSIM, and perceptual quality while requiring fewer sampling steps.

By Davide Evangelista
arXiv Machine Learning
Jul 15

GenDiff: A Dose and Anatomy Aware Diffusion Model with Structural Prior Refinement for Low-Dose CT Reconstruction and Generalization

arXiv:2607. 11941v1 Announce Type: cross Abstract: Computed tomography (CT) is a critical imaging modality for clinical diagnosis, but reducing radiation dose inevitably introduces severe noise and structured artifacts that degrade image quality.

By Md Imam Ahasan, Guangchao Yang, A F M Abdun Noor, Kah Ong Michael Goh, S. M. Hasan Mahmud, Md Mahfuzur Rahman
arXiv AI
Sep 25

A Multimodal 3D Foundation Model for Light Sheet Fluorescence Microscopy Enables Few-Shot Segmentation, Classification, and Deblurring

The paper presents a 3D foundation model for light sheet fluorescence microscopy (LSM) that is pretrained on a large curated set of 3D images from various organisms, stains, and imaging protocols. By jointly optimizing for masked reconstruction and image‑text alignment, the model learns transferable volumetric representations that dramatically reduce the need for annotated data. The pretrained backbone enables efficient few‑shot adaptation to downstream tasks such as segmentation, classification, and deblurring, consistently outperforming baselines according to standard metrics and expert evaluation.

By Adina Scheinfeld, Haotan Zhang, Shang Mu, Rudolf L. M. van Herten, Lucas Stoffl, Ali Erturk, Zhuhao Wu, Johannes C. Paetzold
arXiv Computer Vision
Sep 23

MIAR: Medical Image Super-Resolution With Autoregressive Modeling

MIAR introduces a multi‑scale autoregressive framework for medical image super‑resolution, treating the task as a conditional, progressive next‑scale prediction. It incorporates a Scale‑Adaptive Structural Decoder to preserve structural fidelity and uses a hierarchical beam search during inference to reduce recursive error accumulation. Experiments show MIAR outperforms existing methods, achieving a 7.86% MUSIQ improvement and a 2.02× speedup over diffusion‑based approaches.

By Fang Li, Yinglong Li, Hongyu Wu, Yang Gao, Minwei Zhao, Aimin Hao