arXiv Computer Vision By Hanzhou Liu, Peng Jiang, Jia Huang, Mi Lu

Lumos3D: A Single-Forward Framework for Low-Light 3D Scene Restoration

Read the original on arXiv Computer Vision →

Lumos3D introduces a pose‑free, single‑forward framework for restoring 3D scenes captured in low light. It employs a cross‑illumination distillation scheme where a frozen teacher network provides accurate geometric cues to a student model, and a specialized Lumos loss refines the 3D Gaussian reconstruction. Trained on a single dataset, Lumos3D can directly restore illumination and structure from unposed, low‑light multi‑view images without per‑scene optimization, achieving competitive results on real‑world datasets.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.

Hugging Face Trending Papers
Jul 20

FF-ProCams: Feed-Forward Gaussian Splatting for Projector-Camera System

Projector-camera (ProCams) systems achieve active scene perception and controllable appearance manipulation via structured illumination, serving as a core infrastructure for spatial augmented reality, projection mapping, and surface reflectance acquisition. Existing inverse-rendering methods for ProCams deliver high-fidelity results but rely on time-consuming per-scene optimization, while mainstream feed-forward 3D reconstruction models produce baked appearance that cannot adapt to spatially varying projector illumination.

arXiv Computer Vision
Sep 23

MirrorDistill: Illumination-Aware Latent Distillation for Efficient Low-Light Restoration

MirrorDistill introduces an illumination‑aware latent distillation framework for low‑light image enhancement. It trains a lightweight student encoder‑decoder by aligning its intermediate features with clean‑domain targets generated by a teacher decoder, using feature mirroring and illumination‑aware weighting to emphasize underexposed regions. The method achieves state‑of‑the‑art performance on the LOL‑v2‑Real benchmark while maintaining the lowest computational complexity, and the code is released as open source.

By Farida Mohsen, Tala Zaim, Nurul Izni Rusli, Ali Al-Zawqari, Ali Safa, Samir Brahim Belhaouari
arXiv AI
Sep 10

RelightFormer: Feed-forward Generative Transformer for Multiview Object Relighting

RelightFormer is a feed‑forward generative Transformer that performs single‑ and multi‑view image relighting without explicit intrinsic property estimation. It incorporates a latent illumination module that injects target environment maps into spatial features via cross‑attention, and uses permutation‑invariant positional encodings to process unordered multi‑view inputs symmetrically. Trained on the large Laval Objaverse Dataset, the model achieves state‑of‑the‑art visual and photorealistic relighting quality, and demonstrates strong zero‑shot generalization across various relighting tasks.

By Hejun Wang, Jinxi Li, Junwei Jiang, Shiwei Mao, Hu Cheng, Shouwang Huang, Bo Yang
arXiv Computer Vision
Sep 25

One View Is Enough: In-the-Wild Monocular Pretraining for Novel View Generation

The paper introduces OVIE, a monocular novel-view synthesis method that eliminates the need for multi‑view training data. By using a frozen depth estimator to generate pseudo‑target views from single images and applying masked and adversarial losses, OVIE is trained on 30 million uncurated images. It achieves state‑of‑the‑art performance on RealEstate10K and DL3DV, produces highly consistent multi‑view trajectories, and runs at 116 FPS—over 600× faster than the fastest baseline.

By Adrien Ramanana Rahary, Nicolas Dufour, Patrick Perez, David Picard