Hugging Face Trending Papers

Bricker to BRACE: A Bracket Exposure RAW Dataset and Restoration Model for Flicker-Banding

Flicker-banding (FB), arises from temporal aliasing between a camera's rolling shutter and a display's brightness modulation, degrading screen-captured image readability with color shifts and jagged patterns. Existing single-frame methods with simplified parametric stripe models cannot reliably distinguish these artifacts from genuine texture.

arXiv Computer Vision
1d ago

Seeing Through the Glare: A Multi-Source Benchmark and Ocular-Adaptive Pixel MeanFlow for Eyeglass Reflection Removal

The paper introduces OcuBench, a comprehensive benchmark for eyeglass reflection removal that includes 10,280 synthetic pairs, 732 real-input pseudo-pairs, and 458 real-world test images, enabling both paired evaluation and assessment beyond generated supervision. It also proposes OcuFlow, an ocular-adaptive pixel MeanFlow framework that uses geometry-adaptive representation and one-step pMF to focus on reflection-obscured ocular regions while preserving native-resolution details. Experiments show OcuFlow consistently outperforms baselines in reflection removal quality, ocular fidelity, and efficiency, achieving 67.32% of selections in a blind user study, six times the next-best share.

By Tao Liu, Youwei Pang, Kailai Zhou, Jiaming Zuo, Hanqi Liu, Wei Ji, Peng-Tao Jiang, Xiaofeng Liu, Weisi Lin, Xiaoqi Zhao
arXiv Computer Vision
Sep 10

SloMoDeblur: A Large-Scale Smartphone Image Deblurring Dataset

arXiv:2506.19445v5 Announce Type: replace Abstract: Motion blur remains one of the most common and visually disruptive degradations in real-world smartphone imaging, yet existing deblurring benchmark...

By Syed Mumtahin Mahmud, Mahdi Mohd Hossain Noki, Prothito Shovon Majumder, Abdul Mohaimen Al Radi, Sudipto Das Sukanto, Afia Lubaina, Md. Mosaddek Khan
arXiv Computer Vision
Sep 11

InstantHDR: Single-forward Gaussian Splatting Initialization for HDR 3D Reconstruction

InstantHDR is a feed-forward network that initializes high dynamic range (HDR) 3D scenes from uncalibrated multi-exposure low dynamic range (LDR) image collections in a single forward pass. It uses geometry-guided appearance modeling for multi-exposure fusion and a meta-network for scene-specific tone mapping. The authors also created a pre-training dataset, HDR-Pretrain, with 168 Blender-rendered scenes to support generalizable HDR models, achieving a speedup of about 700× over state‑of‑the‑art optimization methods while maintaining comparable quality after lightweight post‑optimization.

By Dingqiang Ye, Jiacong Xu, Jianglu Ping, Yuxiang Guo, Chao Fan, Vishal M. Patel
arXiv AI
Sep 2

DiffHDR: Re-Exposing LDR Videos with Video Diffusion Models

arXiv:2604.06161v3 Announce Type: replace-cross Abstract: Most digital videos are stored in 8-bit low dynamic range (LDR) formats, where much of the original high dynamic range (HDR) scene radiance i...

By Zhengming Yu, Li Ma, Mingming He, Leo Isikdogan, Yuancheng Xu, Dmitriy Smirnov, Pablo Salamanca, Dao Mi, Pablo Delgado, Ning Yu, Julien Philip, Xin Li, Wenping Wang, Paul Debevec
Hugging Face Trending Papers
Jun 28

EvLIR: Learning Illumination Residuals from Ordered Events for Low-Light Image Enhancement

Low-light image enhancement is severely ill-posed when the input frame contains missing structure, saturated noise, and weak local contrast. Event cameras provide asynchronous brightness-change observations with high temporal resolution, but prior works often treat voxel channels as an unordered or static feature stack before fusion, rather than explicitly modeling their within-window temporal evolution, weakening the temporal evidence that makes events useful.