arXiv Computer Vision

Guided Trajectory Optimization with Sparse Scaling for Test-Time Diffusion

arXiv Computer Vision
Aug 25

Pixel-Space Diffusion via Observation Operators

Pixel‑Space Diffusion via Observation Operators introduces a new framework for pixel‑space diffusion models that addresses a scale‑time mismatch in existing methods. By replacing fixed full‑image supervision with a time‑indexed observation trajectory that progresses from coarse structures to the full image, the model aligns supervision with the natural recovery order of image details. The approach employs Gaussian‑Lanczos operators and a GL‑CoDA decoder to refine features progressively, resulting in faster convergence and higher generation quality, achieving an FID of 1.52 on ImageNet‑256.

By Shaojie Guo, Lichen Ma, Haoyang Tong, Yu He, Zipeng Guo, Xiaoan Liu, Feng Yan, Yu Guo, Fei Wang, Junshi Huang, Yan Wang
arXiv AI
Sep 4

A Posterior-Dynamics Framework for Imaging Inverse Problems with Pretrained Diffusion Priors

The paper introduces a Posterior‑Dynamics Framework that leverages pretrained diffusion models as multiscale priors for linear imaging inverse problems such as deblurring, super‑resolution, and inpainting. By constructing a surrogate likelihood centered on the clean image and incorporating diffusion uncertainty, the authors derive continuous posterior dynamics and a tunable Langevin component for adaptive exploration. They prove theoretical guarantees (endpoint consistency, finite‑horizon tracking, weak accuracy) and present the PD‑IMEX sampler, which achieves high‑quality reconstructions with only 100 score evaluations and controllable fidelity‑diversity trade‑offs.

By Zhaoqiang Liu, Tongyao Pang, Ruibing Wang, Yang Zheng
arXiv AI
Aug 24

Consistency Models for Fast MRI Reconstruction Using Regularization by Denoising

The paper introduces CM-RED, a fast MRI reconstruction method that combines a pretrained consistency model with the regularization by denoising framework. By integrating controlled noise injection into accelerated proximal gradient updates, CM-RED achieves high‑quality reconstructions on fastMRI knee and brain datasets with only four network function evaluations. It consistently outperforms existing diffusion‑ and consistency‑based approaches in quantitative metrics, visual fidelity, and robustness to hyperparameter changes.

By Merve G\"ulle, Junno Yun, Ya\c{s}ar Utku Al\c{c}alar, Mehmet Ak\c{c}akaya
arXiv Computer Vision
Sep 22

AlignMorph: Tuning-Free Diffusion Image Morphing via Explicit Semantic Transport

AlignMorph is a tuning‑free diffusion framework for image morphing that separates geometric alignment from generative denoising. It uses Global Semantic Transport—entropic optimal transport and reliability‑aware latent warping—to achieve diffusion‑compatible semantic alignment, and Coordinate‑Aligned Generation—symmetric bi‑phase attention handoff—to preserve spatial coordinates during denoising. The method eliminates ghosting and delivers superior structural coherence and temporal smoothness on morphing benchmarks without any per‑pair optimization.

By Wuyi Liu, Xu Han, Yuren Chen, Yige Mao, Zishuo Peng, Xianzhi Li