arXiv Computer Vision
Aug 25

Pixel-Space Diffusion via Observation Operators

Pixel‑Space Diffusion via Observation Operators introduces a new framework for pixel‑space diffusion models that addresses a scale‑time mismatch in existing methods. By replacing fixed full‑image supervision with a time‑indexed observation trajectory that progresses from coarse structures to the full image, the model aligns supervision with the natural recovery order of image details. The approach employs Gaussian‑Lanczos operators and a GL‑CoDA decoder to refine features progressively, resulting in faster convergence and higher generation quality, achieving an FID of 1.52 on ImageNet‑256.

By Shaojie Guo, Lichen Ma, Haoyang Tong, Yu He, Zipeng Guo, Xiaoan Liu, Feng Yan, Yu Guo, Fei Wang, Junshi Huang, Yan Wang
arXiv AI
Sep 4

A Posterior-Dynamics Framework for Imaging Inverse Problems with Pretrained Diffusion Priors

The paper introduces a Posterior‑Dynamics Framework that leverages pretrained diffusion models as multiscale priors for linear imaging inverse problems such as deblurring, super‑resolution, and inpainting. By constructing a surrogate likelihood centered on the clean image and incorporating diffusion uncertainty, the authors derive continuous posterior dynamics and a tunable Langevin component for adaptive exploration. They prove theoretical guarantees (endpoint consistency, finite‑horizon tracking, weak accuracy) and present the PD‑IMEX sampler, which achieves high‑quality reconstructions with only 100 score evaluations and controllable fidelity‑diversity trade‑offs.

By Zhaoqiang Liu, Tongyao Pang, Ruibing Wang, Yang Zheng
arXiv AI
Aug 24

Consistency Models for Fast MRI Reconstruction Using Regularization by Denoising

The paper introduces CM-RED, a fast MRI reconstruction method that combines a pretrained consistency model with the regularization by denoising framework. By integrating controlled noise injection into accelerated proximal gradient updates, CM-RED achieves high‑quality reconstructions on fastMRI knee and brain datasets with only four network function evaluations. It consistently outperforms existing diffusion‑ and consistency‑based approaches in quantitative metrics, visual fidelity, and robustness to hyperparameter changes.

By Merve G\"ulle, Junno Yun, Ya\c{s}ar Utku Al\c{c}alar, Mehmet Ak\c{c}akaya