RelightFormer is a feed‑forward generative Transformer that performs single‑ and multi‑view image relighting without explicit intrinsic property estimation. It incorporates a latent illumination module that injects target environment maps into spatial features via cross‑attention, and uses permutation‑invariant positional encodings to process unordered multi‑view inputs symmetrically. Trained on the large Laval Objaverse Dataset, the model achieves state‑of‑the‑art visual and photorealistic relighting quality, and demonstrates strong zero‑shot generalization across various relighting tasks.
By Hejun Wang, Jinxi Li, Junwei Jiang, Shiwei Mao, Hu Cheng, Shouwang Huang, Bo Yang
The paper introduces a benchmark called Shedding Light to evaluate how well generative image models understand and reproduce lighting. The benchmark tests models by asking them to inpaint a simple object, called a light probe, into real photographs and then compares the generated probe to the ground truth to assess lighting direction, colour, and radiance. The authors provide a scalable protocol and open-source code and data for systematic assessment of photometric accuracy in future models.
By Justine Giroux, Jack Oliver Hilliard, Yannick Hold-Geoffroy, Javier Vazquez-Corral, Jean-Fran\c{c}ois Lalonde
arXiv:2502. 07531v5 Announce Type: replace-cross Abstract: Controllable image-to-video (I2V) generation transforms a reference image into a coherent video guided by user-specified control signals.
By Sixiao Zheng, Zimian Peng, Yanpeng Zhou, Yi Zhu, Hang Xu, Xiangru Huang, Yanwei Fu
arXiv:2606. 22574v2 Announce Type: replace-cross Abstract: While synthetic data generation resolves the manual labeling bottleneck in computer vision, minimizing the syn-to-real domain gap requires optimizing rendering variables.
By Hooman Tavakoli Ghinani, Tatjana Legler, Martin Ruskowski
arXiv:2609.28300v1 Announce Type: new
Abstract: Ill-posed inverse problems require priors to constrain the solution space toward plausible outcomes. In inverse rendering, learned priors modeling the...
By Andreea Ardelean, Bernhard Egger
arXiv:2602. 07343v2 Announce Type: replace-cross Abstract: Robust semantic segmentation of road scenes under adverse illumination, lighting, and shadow conditions remain a core challenge for autonomous driving applications.
By Ruturaj Reddy, Hrishav Bakul Barua, Junn Yong Loo, Thanh Thi Nguyen, Ganesh Krishnasamy
WildRelight is the first in-the-wild dataset designed to evaluate single-image relighting models, featuring high-resolution outdoor scenes captured under strictly aligned, temporally varying natural illuminations paired with high-dynamic-range environment maps. The benchmark demonstrates that state-of-the-art models trained on synthetic data suffer severe domain shifts when applied to real-world imagery. Leveraging the dataset’s temporal structure, the authors introduce a physics-guided inference framework combining Diffusion Posterior Sampling with Temporal Sampling-Aware Test-Time Adaptation, enabling synthetic models to self-supervise and align with real-world statistics on-the-fly.
By Lezhong Wang, Mehmet Onurcan Kaya, Siavash Bigdeli, Jeppe Revall Frisvad
GLOW is a Global Illumination‑aware inverse rendering framework for indoor scenes captured with dynamic co‑located light and camera setups. It combines a neural implicit surface representation with a neural radiance cache to jointly optimize geometry and reflectance, while introducing a dynamic radiance cache and a surface‑angle‑weighted radiometric loss to handle near‑field motion, strong inter‑reflections, and specular highlights. Experiments demonstrate that GLOW significantly outperforms prior methods in estimating material reflectance under both natural and co‑located illumination.
By Jiaye Wu, Saeed Hadadan, Geng Lin, Peihan Tu, Auguste Gezalyan, Matthias Zwicker, David Jacobs, Roni Sengupta
The paper introduces ShadowCLR, an unsupervised framework for removing shadows from images without requiring paired shadow–shadow-free data or shadow masks. By leveraging consistency across multiple shadowed observations of the same scene, the method regularizes the model to recover scene-consistent appearance while suppressing shadow-specific variations. Experiments on several benchmarks show that ShadowCLR achieves competitive or superior performance compared to existing unsupervised approaches.
By Anh-Kiet Duong, Petra Gomez-Kr\"amer, Jean-Michel Carozza
arXiv:2608.29925v1 Announce Type: new
Abstract: Controllable image relighting is an important problem in image editing, and hand-drawn scribbles provide an intuitive interface for specifying the desi...
By Xuanpu Zhang, Xuesong Niu, Haoxiang Cao, Ruidong Chen, Jianhao Zeng, Changqian Yu
arXiv:2605.22455v2 Announce Type: replace-cross
Abstract: Real-world deployment of AI vision models is both fueled and limited by the data available for training and testing. Real datasets are sparse...
By Valeria Pais, Malena Mendilaharzu, Daniele Faccio, Luis Oala, Christoph Clausen, Bruno Sanguinetti
arXiv:2604. 23094v2 Announce Type: replace-cross Abstract: Portrait relighting is a low-level vision problem in which physically plausible illumination transfer, identity preservation, and compact real-time inference must be considered together.
By Qian Huang, Mayoore Selvarasa Jaiswal, Zhen Zhong, Rochelle Pereira, Jianyuan Min