arXiv:2608. 09133v1 Announce Type: cross Abstract: Image super-resolution (SR) with large generative models has recently achieved remarkable perceptual quality, yet maintaining fidelity to the LR observation remains challenging.
By Yu Shi, Yuyao Zhang, Yu-wing Tai
arXiv:2608.30782v1 Announce Type: new
Abstract: Real-world image super-resolution (Real-ISR) aims to preserve structures supported by the degraded observation while reconstructing perceptually realis...
By Bingtian Qiao, Yue Shi, Yong Guo, Wenjun Zhang, Jiezhang Cao
SR‑Ground is a large‑scale dataset created to enable fine‑grained segmentation of visual artifacts in super‑resolved images. It contains 63,000 images processed by various state‑of‑the‑art SR models, each annotated at the pixel level for six distinct artifact types, validated through a crowdsourcing study with 1,062 participants. The dataset improves the training of image quality assessment models with grounding capabilities and supports a fine‑tuning pipeline that reduces perceptible artifacts in SR outputs, outperforming no‑reference methods on both benchmark and real‑world low‑resolution datasets.
By Artem Borisov, Evgeney Bogatyrev, Khaled Abud, Dmitriy Vatolin
arXiv:2609.15120v1 Announce Type: new
Abstract: Benefiting from the powerful generative priors of diffusion models, diffusion-based real-world image super-resolution (Real-ISR) methods have demonstra...
By Shuhao Han, Wenjie Liao, Hayden Vance, Hang Dong, Rui Zhang, Chun-Le Guo, Chongyi Li
arXiv:2609.30988v1 Announce Type: new
Abstract: Real-world image super-resolution (SR) requires recovering perceptually realistic high-resolution images from complex low-resolution observations while...
By Xin Di, Mingyu Shi, Yuanfei Bao, Long Peng, Yue Zhao, Jiaming Guo, Renjing Pei, Xueyang Fu, Yang Cao, Zheng-Jun Zha
Diffusion-based generative models have achieved remarkable success in real-world image super-resolution (SR). With tiled diffusion techniques, these models can produce high-resolution images that exceed their native-supported resolution.
arXiv:2608. 15694v1 Announce Type: cross Abstract: Conditional image-to-image generators are single-shot: they map input features to an output in one forward pass and treat it as final, with no opportunity to improve on it.
By Kareem Hassani, Chaymaa Abbas, Hadi Al Mubasher, Mariette Awad
SelfLift is a progressive‑resolution framework that accelerates few‑step diffusion models by enabling late, self‑recovering transitions between low‑ and high‑resolution latents. It introduces a training‑free Artifact‑Aware Consistency Lift that uses disagreement between direct latent lifting and pixel‑VAE re‑encoding to detect and correct artifacts, and a self‑recovery policy that transfers high‑resolution guidance from an internal teacher. Experiments on FLUX.2‑Klein and Z‑Image‑Turbo show latency reductions of 41.5% and 44.1%, and overall speedups of 29.61× and 19.21× over 50‑step baselines while maintaining competitive generation quality.
By Tingyan Wen, Chenqian Yan, Xurui Peng, Xiazhang Fang, Shuai Wang, Xueqian Wang, Songwei Liu
arXiv:2610.07720v1 Announce Type: cross
Abstract: Multi-reference image generation requires preserving the appearance of multiple subjects while composing them into a coherent scene. However, existin...
By Wanning He, Yuyao Zhang, Yu-Wing Tai
arXiv:2603.20186v2 Announce Type: replace
Abstract: In this work, we propose Image-to-Image Rectified Flow Reformulation (I2I-RFR), a practical plug-in reformulation that recasts standard I2I regress...
By Satoshi Iizuka, Shun Okamoto, Kazuhiro Fukui
arXiv:2511. 18050v1 Announce Type: cross Abstract: Diffusion transformers have recently delivered strong text-to-image generation around 1K resolution, but we show that extending them to native 4K across diverse aspect ratios exposes a tightly coupled failure mode spanning positional encoding, VAE compression, and optimization.
By Tian Ye, Song Fei, Lei Zhu
arXiv:2601.17723v3 Announce Type: replace
Abstract: Implicit neural representation (INR) has become the standard approach for arbitrary-scale image super-resolution (ASSR). However, no systematic emp...
By Tayyab Nasir, Daochang Liu, Ajmal Mian