arXiv Computer Vision By Fang Li, Yinglong Li, Hongyu Wu, Yang Gao, Minwei Zhao, Aimin Hao

MIAR: Medical Image Super-Resolution With Autoregressive Modeling

Read the original on arXiv Computer Vision →

MIAR introduces a multi‑scale autoregressive framework for medical image super‑resolution, treating the task as a conditional, progressive next‑scale prediction. It incorporates a Scale‑Adaptive Structural Decoder to preserve structural fidelity and uses a hierarchical beam search during inference to reduce recursive error accumulation. Experiments show MIAR outperforms existing methods, achieving a 7.86% MUSIQ improvement and a 2.02× speedup over diffusion‑based approaches.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.

arXiv Machine Learning
Sep 3

Perceptually Regularized Diffusion Model for Image Super-Resolution

The paper introduces a perceptually regularized diffusion framework for image super‑resolution, adding perceptual‑loss based regularization to the standard diffusion training objective. This approach incorporates prior knowledge to improve training convergence and encourages the recovery of meaningful image features. Experiments on benchmark datasets show enhanced perceptual quality while maintaining competitive distortion metrics.

By Chuxiangbo Wang, Pavithra Venkatachalapathy, Ying Liang, Min Wang, Jing Qin, Yifei Lou, Weihong Guo
arXiv Computer Vision
Sep 2

Prior-Guided Implicit Neural Representations for Single-Subject Diffusion MRI Super-Resolution

The paper introduces a transfer‑learning framework that pre‑trains an implicit neural representation (INR) on a high‑resolution diffusion MRI template and then adapts it to individual subjects through registration and fine‑tuning. This approach enables native single‑subject super‑resolution, achieving a 4× through‑plane up‑sampling from 5 mm to 1.25 mm on Human Connectome Project data. Compared to a recent baseline, the method reduces NRMSE by 36–49 % and increases FSIM by 24–43 %, while training 6× faster and outperforming other INR‑based techniques on both image quality and domain‑specific metrics.

By Abdulkader Ghandoura, Marsil Zakour, William Consagra, Yogesh Rathi