arXiv Computer Vision By SaiKiran Tedla, Francesco Banterle, Trevor Canham, Karanpreet Raja, David B. Lindell, Kiriakos N. Kutulakos, Jiacheng Li, Feiran Li, Daisuke Iso

Generating HDR Video from SDR Video

Read the original on arXiv Computer Vision →

The paper presents a framework for converting standard dynamic range (SDR) videos into high dynamic range (HDR) videos using large-scale generative video models. It introduces a Multi-Exposure Video Model (MEVM) that predicts exposure-bracketed linear SDR sequences from a single nonlinear SDR input, and a Video Merging Model (VMM) that fuses these predictions into a high-quality HDR sequence while preserving detail in shadows and highlights. Experiments, qualitative evaluation, and a user study demonstrate robust HDR conversion for casual consumer footage and iconic films, and the approach can be integrated into existing SDR generative video pipelines.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.

arXiv AI
Sep 2

DiffHDR: Re-Exposing LDR Videos with Video Diffusion Models

arXiv:2604.06161v3 Announce Type: replace-cross Abstract: Most digital videos are stored in 8-bit low dynamic range (LDR) formats, where much of the original high dynamic range (HDR) scene radiance i...

By Zhengming Yu, Li Ma, Mingming He, Leo Isikdogan, Yuancheng Xu, Dmitriy Smirnov, Pablo Salamanca, Dao Mi, Pablo Delgado, Ning Yu, Julien Philip, Xin Li, Wenping Wang, Paul Debevec
arXiv Computer Vision
5d ago

Recurrent Dynamic Range Extension

The paper introduces a method for progressively extending the dynamic range of an image by learning to increase it by a single exposure value first, then applying the network recurrently to achieve full HDR reconstruction. The approach is agnostic to input dynamic range, targets a bounded output domain, and utilizes RAW images with adversarial losses to produce realistic results. Memory Replay during backpropagation allows training over multiple inference stages, reducing reconstruction errors and enabling robust recovery of bright highlights in long‑tailed HDR scenes.

By Sebastian Dille, Keru Fu, S. Mahdi H. Miangoleh, Ya\u{g}{\i}z Aksoy
arXiv Computer Vision
Sep 11

InstantHDR: Single-forward Gaussian Splatting Initialization for HDR 3D Reconstruction

InstantHDR is a feed-forward network that initializes high dynamic range (HDR) 3D scenes from uncalibrated multi-exposure low dynamic range (LDR) image collections in a single forward pass. It uses geometry-guided appearance modeling for multi-exposure fusion and a meta-network for scene-specific tone mapping. The authors also created a pre-training dataset, HDR-Pretrain, with 168 Blender-rendered scenes to support generalizable HDR models, achieving a speedup of about 700× over state‑of‑the‑art optimization methods while maintaining comparable quality after lightweight post‑optimization.

By Dingqiang Ye, Jiacong Xu, Jianglu Ping, Yuxiang Guo, Chao Fan, Vishal M. Patel
arXiv Machine Learning
Jun 9

MilliVid: Hierarchical Latents for Long-Range Consistency in Video Generation

arXiv:2606. 09056v1 Announce Type: cross Abstract: Video generative models have become increasingly powerful, but long-range consistency remains challenging to achieve because even a few dozen frames require impractically long transformer sequence lengths.

By Ishaan Preetam Chandratreya, David Charatan, Basile Van Hoorick, Sergey Zakharov, Vitor Guizilini, Phillip Isola, Vincent Sitzmann