arXiv:2606. 17584v1 Announce Type: cross Abstract: Finding the initial noise that generates a given data sample, known as inversion, is a key component for downstream applications such as training-free image editing.
By Semin Kim, Jihwan Yoon, Seunghoon Hong
arXiv:2605. 16399v2 Announce Type: replace-cross Abstract: The inversion of diffusion models plays a central role in image editing.
By Barbora Barancikova, Daniil Shmelev, Cristopher Salvi
arXiv:2608.22773v1 Announce Type: new
Abstract: Dynamic 3D Gaussian Splatting (3DGS) achieves photorealistic reconstruction of time-varying scenes, and recent physics-aware extensions improve extrapo...
By Shogo Sato, Takuhiro Kaneko, Shoichiro Takeda, Tomoyasu Shimada, Riku Inoue, Kazuhiko Murasaki, Ryuichi Tanida
FlashRender is a few-step generative rendering framework that quickly retakes a source video along a target camera trajectory. It addresses discretization error by introducing Representation Transformation and Alignment (RETA) to align source-video representations with target-video features, reducing denoising trajectory curvature. The model is further refined with a MeanFlow objective and on-policy flow map distillation, achieving video quality and geometric consistency comparable to multi-step baselines at a 25× lower sampling cost while improving camera controllability.
By Byeongjun Park, Byung-Hoon Kim, Hyungjin Chung
Geometric foundation models, such as the Visual Geometry Grounded Transformer (VGGT), provide strong 3D priors from unposed images. However, such models operate purely in a feed-forward, deterministic regime, \ie~they cannot generate plausible geometry beyond what the input views directly support.
arXiv:2602. 10099v2 Announce Type: replace Abstract: Leveraging representation encoders for generative modeling offers a path for efficient, high-fidelity synthesis.
By Amandeep Kumar, Vishal M. Patel
Bi-FlowGS introduces a bidirectional co-refinement framework that links generative view completion with 3D Gaussian Splatting geometry. It employs Video-to-Geometry Flow Distillation (V2G) to transfer temporal correspondence from restored videos into Gaussian geometry, mitigating the Geometry Cheating problem. Simultaneously, Geometry-to-Video Flow-Guided Restoration (G2V) uses the current 3DGS geometry to guide temporally consistent video restoration, creating a loop where restored videos and optimized geometry iteratively improve each other, leading to better rendering quality and geometric consistency on wide-baseline and 360° benchmarks.
By Yuetong Wang, Jinsheng Quan, Yi Yang, Yawei Luo
arXiv:2609.08084v1 Announce Type: cross
Abstract: Monocular depth estimation is a ubiquitous yet highly ill-posed computer vision task, with downstream applications in scene reconstruction, computati...
By Igor Pavlovic, Thiemo Wandel, Anton Obukhov, Luca Bartolomei, Andrey Davydov, Fabio Tosi, Matteo Poggi, Sabine S\"usstrunk, Dengxin Dai
The paper introduces Curvature-Adaptive Tubular Correction (CAT), a training‑free plugin that refines diffusion guidance by decomposing the guidance gradient into normal and tangent components and regulating them within a noise‑dependent geometric budget. CAT charges normal displacement at first order and tangent displacement according to directional curvature, solving a one‑dimensional dual equation for optimal magnitudes and using Armijo backtracking to calibrate the step size. Experiments on seven inverse problems with FFHQ and ImageNet demonstrate that CAT consistently improves pixel‑ and latent‑space samplers, enhances perceptual metrics, and achieves the lowest FID across classifier‑free guidance scales while maintaining stable saturation and contrast.
By Enze Jiang, Jinwei He, Zheng Ma
arXiv:2601. 19180v2 Announce Type: replace-cross Abstract: Inversion-free image editing using flow-based generative models challenges the prevailing inversion-based pipelines.
By Lifan Jiang, Boxi Wu, Yuhang Pei, Tianrun Wu, Yongyuan Chen, Yan Zhao, Shiyu Yu, Deng Cai
arXiv:2606. 19802v1 Announce Type: new Abstract: Image restoration faces a fundamental tradeoff: methods that minimize error produce blurry reconstructions, while those that maximize perceptual quality yield sharp but less faithful images.
By Nicolas Zilberstein, Morteza Mardani, Santiago Segarra
EquiReg introduces an equivariance‑regularized diffusion framework that penalises sampling trajectories deviating from the data manifold, thereby improving posterior sampling for inverse problems. By formalising manifold‑preferential equivariant functions—naturally arising from data augmentation or inherent symmetries—EquiReg guides diffusion steps toward symmetry‑preserving regions of the solution space. The method shows consistent gains in both linear and nonlinear image restoration tasks and partial differential equation solving, especially under reduced sampling and measurement consistency steps, and is available as open‑source code.
By Bahareh Tolooshams, Aditi Chandrashekar, Rayhan Zirvi, Abbas Mammadov, Jiachen Yao, Chuwei Wang, Anima Anandkumar