We present Generative Relightable Avatars (GRA), a person-specific method for photorealistic free-view rendering and environment-map relighting of full-body humans. We postulate that modeling fine-grained appearance details is inherently a one-to-many problem that can benefit from a generative formulation.
arXiv:2606. 31637v1 Announce Type: cross Abstract: Intrinsic decomposition which expresses image colors as the product of diffuse albedo and shading, possibly augmented with view-dependent residuals has a long history in image editing as it enables the modification of object colors and textures without altering lighting.
By Alexandre Lanvin, Jeffrey Hu, Simon Lucas, Adrien Bousseau, George Drettakis
arXiv:2604. 10359v3 Announce Type: replace-cross Abstract: Low-light image enhancement (LLIE) aims to restore natural visibility, color fidelity, and structural detail under severe illumination degradation.
By Alexandru Brateanu, Tingting Mu, Codruta Ancuti, Cosmin Ancuti
arXiv:2606. 06899v1 Announce Type: cross Abstract: Variations in illumination remain a major challenge for visual representation learning, as they induce substantial appearance changes both across and within environments.
By Lizhen Zhu, Charantej Reddy Pochimireddy, James Z Wang, Brad Wyble
arXiv:2608. 10798v1 Announce Type: cross Abstract: Most image colorization systems operate in $Lab$ space by predicting chroma ($ab$) while preserving an input-derived luminance channel ($L$).
By Swarnim Maheshwari, Syed Imam Ali, Vineeth N. Balasubramanian
Large-scale text-to-image models are attractive backbones for dense prediction because RGB generation pretraining learns rich semantic, structural, and geometric priors. Existing generative and editing approaches reuse these priors by casting dense prediction as target generation: annotations such as depth, normals, alpha mattes, masks, and heatmaps are encoded into an RGB-trained VAE latent space and decoded back as image-like targets.
arXiv:2608.29269v1 Announce Type: new
Abstract: Relightable interactive scene reconstruction aims to build an editable 3D model from scans of different object arrangements and render new layouts unde...
By Haonan Zhou, Gaoxiang Linghu, Youlin Jia, Hongyu Cui, Kewei Wei, Kaiyue Zhou, Bruce X. B. Yu, Gaoang Wang
arXiv:2606. 29379v1 Announce Type: cross Abstract: Gaussian splatting (GS) has garnered significant attention in VR/AR and digital content creation due to its explicit parameterization and efficient rendering capabilities.
By Jiaxin Li, Tong Wu, Yi Wei, Tailin Wu, Li Zhang
Luce is a 3D representation that unifies geometry and physically based rendering (PBR) materials within a voxelized multimodal Gaussian cloud, using dedicated Gaussian primitives for each modality. A variational autoencoder compresses this representation into a unified material‑aware latent space, which a rectified‑flow transformer generates from a single image conditioned on multi‑layer features from a pretrained image encoder. The latent decodes into relightable PBR Gaussians and an optional textured mesh with a tangent‑space normal map, achieving state‑of‑the‑art single‑image‑to‑3D generation on Toys4K and improving CLIP image‑alignment scores on a benchmark of AI‑generated images.
By Mayank Singh, Michele Stoppa, Alvise Memo, Rui Yu, Harsha Kalli, Srimanth Gunturi, Muhammad Ahmed Riaz, Behrooz Shahsavari, Waleed Abdulla, David E. Jacobs
arXiv:2502. 07531v5 Announce Type: replace-cross Abstract: Controllable image-to-video (I2V) generation transforms a reference image into a coherent video guided by user-specified control signals.
By Sixiao Zheng, Zimian Peng, Yanpeng Zhou, Yi Zhu, Hang Xu, Xiangru Huang, Yanwei Fu
arXiv:2607. 08185v1 Announce Type: cross Abstract: Enhancing images to make them visually appealing is a persistent challenge in computer vision.
By David Serrano-Lozano, Luis Herranz, Michael S. Brown, Javier Vazquez-Corral
arXiv:2609.00901v1 Announce Type: new
Abstract: Modifying the illumination of driving images is a fundamental challenge, as most datasets are captured at specific times of day. Existing methods rely...
By Hala Djeghim, Nathan Piasco, Luis Rold\~ao, Moussab Bennehar, Dzmitry Tsishkou, C\'eline Loscos, D\'esir\'e Sidib\'e