arXiv:2606. 31290v1 Announce Type: new Abstract: Diffusion models enable probabilistic super-resolution and conditional generation, but pixel-space methods are computationally expensive and learned latent spaces often lack interpretable uncertainty quantification.
By Onkar Jadhav, Tim French, Matthew Rayson, Nicole L. Jones
arXiv:2606.06176v3 Announce Type: replace
Abstract: Underwater Image Enhancement (UIE) is essential for mitigating degradations caused by water medium. Although learning-based methods have advanced s...
By Haochen Hu, Yanrui Bin, Chih-yung Wen, Bing Wang
arXiv:2606. 02310v1 Announce Type: cross Abstract: Flooding is the most pervasive natural disaster worldwide.
By Yogesh Bhattarai, Vijay Chaudhary, Wai Lim Kim, Sanjib Sharma
Unified image restoration (UIR) aims to recover high-quality (HQ) content from low-quality (LQ) images with different degradations using a single model. Most recent methods adapt large pretrained text-to-image (T2I) latent diffusion models for their strong capacity and generative priors.
arXiv:2608. 08965v1 Announce Type: new Abstract: Underwater images often suffer from diverse and coexisting degradations, including color distortion, scattering haze, texture attenuation, and uneven illumination.
By Weifeng Kong, Chenghao Xu, Lin Chen, Ziheng Cao, Guanying Huo
arXiv:2608. 01298v1 Announce Type: cross Abstract: Diffusion Transformers (DiTs) have emerged as a core architecture in generative modeling due to their scalability and adaptability to multimodal tasks.
By Junno Yun, Ya\c{s}ar Utku Al\c{c}alar, Mehmet Ak\c{c}akaya
arXiv:2606. 30934v1 Announce Type: new Abstract: Modern text-to-image diffusion models, such as diffusion transformers (DiT), rely on timestep or prompt embeddings to modulate the strength of the denoising process in each timestep.
By Luke Budny, Yuhong Guo, Kevin Cheung
arXiv:2606. 04299v1 Announce Type: cross Abstract: We consider the problem of generating images whose internal structure -- defined by the distribution of patches across multiple scales -- matches that of a single reference image.
By Haojun Qiu, Kiriakos N. Kutulakos, David B. Lindell
WaterClear-GS introduces a physics-informed Gaussian splatting method tailored for underwater 3D reconstruction and appearance restoration. It models underwater degradation as intrinsic Gaussian attributes and employs a dual-branch optimization that separates clean appearance from degradation while preserving photometric consistency. The approach incorporates depth-guided geometry regularization, perception-driven supervision, exposure constraints, adaptive regularization, and spectral regularization, achieving strong novel view synthesis and image restoration performance at over 160 FPS.
By Xinrui Zhang, Yufeng Wang, Zesheng Wang, Dacheng Qi, Wenrui Ding, Shuangkang Fang
arXiv:2608. 03822v1 Announce Type: cross Abstract: Developing robust flood assessment models requires high-quality paired satellite imagery, yet such data remain scarce for flood-specific image generation.
By Zhang Weihui, Wang Ruizhi, Xu Hongye, Wang Huiqiong, Sun Li, Song Mingli
arXiv:2609.08084v1 Announce Type: cross
Abstract: Monocular depth estimation is a ubiquitous yet highly ill-posed computer vision task, with downstream applications in scene reconstruction, computati...
By Igor Pavlovic, Thiemo Wandel, Anton Obukhov, Luca Bartolomei, Andrey Davydov, Fabio Tosi, Matteo Poggi, Sabine S\"usstrunk, Dengxin Dai
arXiv:2609.37080v1 Announce Type: new
Abstract: Latent Diffusion Models (LDMs) typically adopt a two-stage pipeline: an auto-encoder (AE) is first pre-trained to define a latent space, then a diffusi...
By Zhengqiang Zhang, Lingchen Sun, Rongyuan Wu, Qiaosi Yi, Xiangtao Kong, Chaodong Xiao, Lei Zhang