Real-world image super-resolution (Real-ISR) aims to reconstruct high-quality (HQ) images from low-quality (LQ) inputs subject to diverse real-world degradations. Recent advances have leveraged the LQ inputs and natural image priors learned by Stable Diffusion models to achieve impressive results.
arXiv:2608.30129v1 Announce Type: new
Abstract: This work presents $\textbf{Lapis}$, a $\textbf{l}$inear-$\textbf{a}$ttention-based $\textbf{pi}$xel-$\textbf{s}$pace generative framework that achieve...
By Bingde Liu, Wu Ran, Jinglei Zhang, Huanhuan Yuan, Chao Ma
arXiv:2608.21786v2 Announce Type: replace
Abstract: General image fusion aims to integrate complementary information from multiple source images, but existing methods often rely on task-specific mode...
By Xingxin Xu, Siqi Zhao, Xin Li, Xinjie Yao, Yiming Sun, Pengfei Zhu
arXiv:2609.00798v1 Announce Type: new
Abstract: Pixel-space diffusion has recently emerged as a promising direction for high-fidelity image generation by modeling images directly in the original pixe...
By Weiyi You, Jinhua Zhang, Xingyu Zhou, Wei Long, Junyu Lou, Shuhang Gu
arXiv:2609.01123v1 Announce Type: new
Abstract: Recent advancements in low-light image enhancement have leveraged diffusion models for their strong ability to generate perceptually realistic, detaile...
By Ruoyu Guo, Haonan Zhong, Maurice Pagnucco, Yang Song
arXiv:2512. 04390v2 Announce Type: replace-cross Abstract: Joint video super-resolution and deblurring (VSRDB) requires both efficient long-range temporal modeling and robustness to frame-wise exposure-duration variation, which changes the extent of motion blur across video frames.
By Geunhyuk Youk, Jihyong Oh, Munchurl Kim
The paper introduces a zero‑shot video restoration and enhancement framework that leverages a text‑to‑image latent diffusion model along with multi‑modal references. It employs dual prompt tuning inversion and sampling to cut inference time to about one‑third of the original, while also strengthening performance and temporal consistency. Additional techniques such as texture‑aware video token merging, referenced self‑attention, and referenced token merging further improve temporal coherence across frames.
By Cong Cao, Huanjing Yue, Xin Liu, Jingyu Yang
Pixel-space diffusion has recently emerged as a promising direction for high-fidelity image generation by modeling images directly in the original pixel domain. However, pixel-space diffusion is compu...
arXiv:2607. 10140v1 Announce Type: cross Abstract: Existing optical flow methods broadly follow two paradigms: iterative optimization and diffusion-based estimation.
By Yuang Meng, Chenyang Wu, Xianshun Liu, Chun-Le Guo, Zichen Liang, Lina Lei, Jie Liang, Hui Zeng, Chongyi Li, Lei Zhang
arXiv:2607. 25275v1 Announce Type: cross Abstract: Real-world Image Restoration (Real-IR) aims to recover high-quality (HQ) images from complex and unknown degradations.
By Zhenning Shi, Chen Xu, Junhao Zhang, Kefei Zhang, Linjie Liu, Zhedong Zheng, Tao Li
arXiv:2604. 10359v3 Announce Type: replace-cross Abstract: Low-light image enhancement (LLIE) aims to restore natural visibility, color fidelity, and structural detail under severe illumination degradation.
By Alexandru Brateanu, Tingting Mu, Codruta Ancuti, Cosmin Ancuti
Recent advancements in low-light image enhancement have leveraged diffusion models for their strong ability to generate perceptually realistic, detailed images. Patch diffusion models further offer a...