arXiv AI

Image-Conditional Diffusion Transformer for Underwater Image Enhancement

arXiv Machine Learning
Jul 1

Patch-PODiff-ViT: Structured Latent Diffusion with Patchwise POD for Super-Resolution and Uncertainty Quantification

arXiv:2606. 31290v1 Announce Type: new Abstract: Diffusion models enable probabilistic super-resolution and conditional generation, but pixel-space methods are computationally expensive and learned latent spaces often lack interpretable uncertainty quantification.

By Onkar Jadhav, Tim French, Matthew Rayson, Nicole L. Jones
arXiv Computer Vision
Sep 25

WaterClear-GS: Optical-Aware Gaussian Splatting for Underwater Reconstruction and Restoration

WaterClear-GS introduces a physics-informed Gaussian splatting method tailored for underwater 3D reconstruction and appearance restoration. It models underwater degradation as intrinsic Gaussian attributes and employs a dual-branch optimization that separates clean appearance from degradation while preserving photometric consistency. The approach incorporates depth-guided geometry regularization, perception-driven supervision, exposure constraints, adaptive regularization, and spectral regularization, achieving strong novel view synthesis and image restoration performance at over 160 FPS.

By Xinrui Zhang, Yufeng Wang, Zesheng Wang, Dacheng Qi, Wenrui Ding, Shuangkang Fang
arXiv Machine Learning
Sep 10

Marigold V2: Revisiting Diffusion Transformers for Monocular Depth Estimation

arXiv:2609.08084v1 Announce Type: cross Abstract: Monocular depth estimation is a ubiquitous yet highly ill-posed computer vision task, with downstream applications in scene reconstruction, computati...

By Igor Pavlovic, Thiemo Wandel, Anton Obukhov, Luca Bartolomei, Andrey Davydov, Fabio Tosi, Matteo Poggi, Sabine S\"usstrunk, Dengxin Dai