The SEE Challenge 2026 invites participants to restore RGB images using synchronized event camera data and a target brightness statistic across a wide illumination range. Using the SEE-600K dataset of 610,126 image‑event pairs from 202 real‑world scenes, teams compete under an open‑system protocol, with PSNR as the primary ranking metric and SSIM as a secondary measure. Fifteen valid submissions were evaluated, revealing closely spaced top scores and consistent local errors under severe underexposure, while the report also examines exposure subsets, semantic test cases, shared failure patterns, and system design choices.
By Yunfan Lu, Mingchao Xu, Hanyu Zhou, Shaoyu Liu, Haoyue Liu, Peiqi Duan, Shihan Peng, Yinqiang Zheng, Boxin Shi, Gim Hee Lee, Hui Xiong, Davide Scaramuzza
IDM-Net is a lightweight Illumination-Decoupled Modulation Network designed for low-light image enhancement. It uses a dual-encoder architecture: a structure encoder extracts multi-scale appearance features from the RGB image, while a lightweight illumination encoder learns illumination priors from the decoupled luminance channel. An Illumination-Guided Modulation module injects these priors into the decoder via spatially adaptive affine modulation, and a Feature Refinement Block progressively suppresses artifacts and recovers fine details, achieving competitive performance on standard benchmarks with good balance between quality and efficiency.
By Cheng-Yen Hsiao, Jing-Ming Guo
arXiv:2609.28300v1 Announce Type: new
Abstract: Ill-posed inverse problems require priors to constrain the solution space toward plausible outcomes. In inverse rendering, learned priors modeling the...
By Andreea Ardelean, Bernhard Egger
arXiv:2602.19202v3 Announce Type: replace
Abstract: Event cameras excel at high-speed, low-power, and high-dynamic-range scene perception. However, as they fundamentally record only relative intensit...
By Gang Xu, Zhiyu Zhu, Junhui Hou
arXiv:2607. 09114v1 Announce Type: cross Abstract: Video anomaly detection (VAD) is critical for automated surveillance but remains fragile under challenging conditions such as illumination variations, fast motion, and complex backgrounds when relying solely on visible light videos.
By Peipei Zhu, Yueqing Niu, Lin Zhu, Guanchong Niu, Yang Yu, Zheng Li
arXiv:2606. 13580v1 Announce Type: cross Abstract: Event-based vision has drawn increasing attention owing to its distinctive properties, including ultra-high temporal resolution and extreme dynamic range.
By Dachun Kai, Jiayao Lu, Yueyi Zhang, Xiaoyan Sun
arXiv:2609.25803v1 Announce Type: new
Abstract: High-rate dense perception in dynamic environments is limited by the low update rate of RGB cameras, as rapid scene changes can occur between frames. E...
By Tao Wan, Xiaoshan Wu, Yifei Yu, Bo Wang, Xiaoyang Lyu, Muxin Liu, Aoxuan Pan, Zhongrui Wang, Xiaojuan Qi
arXiv:2605. 00271v3 Announce Type: replace-cross Abstract: Event cameras provide several unique advantages over standard frame-based sensors, including high temporal resolution, low latency, and robustness to extreme lighting.
By Vincenzo Polizzi, David B. Lindell, Jonathan Kelly
arXiv:2608.29043v1 Announce Type: new
Abstract: Light-effect contamination poses a significant challenge to nighttime visibility enhancement. Most methods suppress light effects by estimating and dec...
By Hanting Li, Xin Sun, Wei Ye, Jungong Han, Liang-jie Zhang
arXiv:2507.01927v3 Announce Type: replace
Abstract: While CNNs and ViTs dominate vision architectures, all-MLP models offer a structurally simpler alternative whose patch-independent processing is na...
By Zhentan Zheng
arXiv:2505. 08438v4 Announce Type: replace-cross Abstract: Event cameras are rapidly emerging as powerful vision sensors for 3D reconstruction, uniquely capable of asynchronously capturing per-pixel brightness changes.
By Chuanzhi Xu, Haoxian Zhou, Langyi Chen, Haodong Chen, Zeke Zexi Hu, Zhicheng Lu, Ying Zhou, Vera Chung, Qiang Qu, Weidong Cai
arXiv:2512. 04390v2 Announce Type: replace-cross Abstract: Joint video super-resolution and deblurring (VSRDB) requires both efficient long-range temporal modeling and robustness to frame-wise exposure-duration variation, which changes the extent of motion blur across video frames.
By Geunhyuk Youk, Jihyong Oh, Munchurl Kim