arXiv Machine Learning

FusionRelight: Relighting Portraits in Real Time via Hybrid Domain Knowledge Fusion

arXiv:2604. 23094v2 Announce Type: replace-cross Abstract: Portrait relighting is a low-level vision problem in which physically plausible illumination transfer, identity preservation, and compact real-time inference must be considered together.

arXiv Computer Vision
Sep 23

MirrorDistill: Illumination-Aware Latent Distillation for Efficient Low-Light Restoration

MirrorDistill introduces an illumination‑aware latent distillation framework for low‑light image enhancement. It trains a lightweight student encoder‑decoder by aligning its intermediate features with clean‑domain targets generated by a teacher decoder, using feature mirroring and illumination‑aware weighting to emphasize underexposed regions. The method achieves state‑of‑the‑art performance on the LOL‑v2‑Real benchmark while maintaining the lowest computational complexity, and the code is released as open source.

By Farida Mohsen, Tala Zaim, Nurul Izni Rusli, Ali Al-Zawqari, Ali Safa, Samir Brahim Belhaouari
arXiv Computer Vision
Aug 27

Lumos3D: A Single-Forward Framework for Low-Light 3D Scene Restoration

Lumos3D introduces a pose‑free, single‑forward framework for restoring 3D scenes captured in low light. It employs a cross‑illumination distillation scheme where a frozen teacher network provides accurate geometric cues to a student model, and a specialized Lumos loss refines the 3D Gaussian reconstruction. Trained on a single dataset, Lumos3D can directly restore illumination and structure from unposed, low‑light multi‑view images without per‑scene optimization, achieving competitive results on real‑world datasets.

By Hanzhou Liu, Peng Jiang, Jia Huang, Mi Lu
Hugging Face Trending Papers
Aug 13

P2Fusion: Prompt-based Progressive Infrared-Visible Image Fusion via Dual-Prior Distillation

Infrared-visible image fusion (IVIF) is pivotal for multimodal perception, yet reconciling the inherent information disparity between thermal and textural features remains a fundamental challenge. Existing prior-guided methods often rely on static constraints that induce optimization conflicts or utilize extrinsic semantic priors from large-scale foundation models (e.

arXiv AI
Sep 10

WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting

WildRelight is the first in-the-wild dataset designed to evaluate single-image relighting models, featuring high-resolution outdoor scenes captured under strictly aligned, temporally varying natural illuminations paired with high-dynamic-range environment maps. The benchmark demonstrates that state-of-the-art models trained on synthetic data suffer severe domain shifts when applied to real-world imagery. Leveraging the dataset’s temporal structure, the authors introduce a physics-guided inference framework combining Diffusion Posterior Sampling with Temporal Sampling-Aware Test-Time Adaptation, enabling synthetic models to self-supervise and align with real-world statistics on-the-fly.

By Lezhong Wang, Mehmet Onurcan Kaya, Siavash Bigdeli, Jeppe Revall Frisvad
Hugging Face Trending Papers
Jul 20

FF-ProCams: Feed-Forward Gaussian Splatting for Projector-Camera System

Projector-camera (ProCams) systems achieve active scene perception and controllable appearance manipulation via structured illumination, serving as a core infrastructure for spatial augmented reality, projection mapping, and surface reflectance acquisition. Existing inverse-rendering methods for ProCams deliver high-fidelity results but rely on time-consuming per-scene optimization, while mainstream feed-forward 3D reconstruction models produce baked appearance that cannot adapt to spatially varying projector illumination.

arXiv AI
Sep 10

RelightFormer: Feed-forward Generative Transformer for Multiview Object Relighting

RelightFormer is a feed‑forward generative Transformer that performs single‑ and multi‑view image relighting without explicit intrinsic property estimation. It incorporates a latent illumination module that injects target environment maps into spatial features via cross‑attention, and uses permutation‑invariant positional encodings to process unordered multi‑view inputs symmetrically. Trained on the large Laval Objaverse Dataset, the model achieves state‑of‑the‑art visual and photorealistic relighting quality, and demonstrates strong zero‑shot generalization across various relighting tasks.

By Hejun Wang, Jinxi Li, Junwei Jiang, Shiwei Mao, Hu Cheng, Shouwang Huang, Bo Yang
arXiv AI
3d ago

After a Decade: Bringing Shadow Removal into the Real World with Agentic Training Data

The paper introduces AgenticShadow, a new dataset of 17,138 image‑mask‑target triplets created through an offline agentic workflow that combines physics‑motivated generation, failure detection, feedback‑driven retry, candidate selection, and deterministic correction. This approach addresses the long‑standing lack of diverse paired shadow‑free training data by leveraging existing shadow detection datasets and producing realistic shadow‑free targets. Models trained on AgenticShadow show significant improvements, reducing color distribution differences by 50.5% and cross‑domain LAB RMSE by 19.7‑37.5% compared to prior work.

By Shilin Hu, Jingyi Xu, Dimitris Samaras, Hieu Le