arXiv:2609.27149v1 Announce Type: new
Abstract: Remote sensing change detection requires both global reasoning across bitemporal images and precise localization of changed regions. However, dense att...
By Anuvab Sen, Maneet Chatterjee, Aparup Ghosh, Udayon Sen, Arnav Aditya, Yixin Zhang
arXiv:2605. 15375v2 Announce Type: replace-cross Abstract: Remote sensing change detection (RSCD) localises changes between two images of the same geographic region.
By Bla\v{z} Rolih, Matic Fu\v{c}ka, Filip Wolf, Luka \v{C}ehovin Zajc
Remote sensing change detection for real-world monitoring often relies on imperfect heterogeneous observations, where pre- and post-event images may be asynchronous, cross-sensor, or affected by illumination, seasonal, and modality shifts. This setting is especially challenging for EO-SAR disaster mapping, where nuisance variation can resemble structural damage.
The core challenge of heterogeneous change detection in remote sensing imagery lies in effectively decoupling genuine land-cover changes from significant modal disparities caused by distinct imaging mechanisms. These intrinsic inconsistencies are prone to introducing pseudo-changes, thereby constraining detection accuracy.
MambaMPD is a new segmentation framework that leverages Vision Mamba models for marine pollution detection in remote‑sensing imagery. It introduces two structural priors—Frequency‑Aware Augmentation (FAA) and multi‑scale Edge‑Guided Attention (EGA)—to better capture low‑contrast, fragmented pollution patterns and sharpen boundaries. Experiments on the MADOS and M4D datasets show that MambaMPD outperforms existing methods in mIoU while using far less computation than foundation‑model approaches.
By Shuaiyu Chen, Wei Han, Peng Ren, Chunbo Luo, Zeyu Fu
The paper introduces MaSoN, an end-to-end unsupervised remote sensing change detection framework that synthesises diverse changes directly in latent feature space during training. By generating changes based on feature statistics of the target data, MaSoN produces data‑driven variations that align with the target domain and can be applied to new modalities such as SAR and multispectral imagery. The method achieves a 14.1 percentage point improvement in average F1 score across five benchmarks, demonstrating strong generalisation across diverse change types.
By Bla\v{z} Rolih, Matic Fu\v{c}ka, Filip Wolf, Luka \v{C}ehovin Zajc
arXiv:2608. 15647v1 Announce Type: cross Abstract: Semantic segmentation of very-high-resolution (VHR) remote sensing imagery increasingly benefits from strong pretrained hierarchical encoders, yet exploiting their multi-stage representations remains difficult.
By Shuaishuai Cao, Meng Tang, Shuwei Peng, Xuan Liu, Min Huang, Jie Chen, Jiacheng Niu, Yong Chen, Edore Akpokodje, Hui Lin
arXiv:2606.20032v2 Announce Type: replace
Abstract: Unlike traditional remote sensing change detection that relies on predefined categories, Open-Vocabulary Change Detection (OVCD) identifies land co...
By Hongming Zhu, Huaji Chen, Bowen Du, Sicong Liu, Qin Liu
arXiv:2608.29626v1 Announce Type: new
Abstract: Salient object detection in optical remote sensing images (ORSI-SOD) requires dense predictions that preserve object completeness and structural contin...
By Yi Xu, Ruichao Hou, Tongwei Ren, Gangshan Wu
The paper proposes a weakly supervised remote sensing change detection method that uses change captions as the sole supervision signal, eliminating the need for pixel‑level change masks. It introduces a caption‑driven generation pipeline to create bi‑temporal image pairs with controlled changes and a Semantic‑Appearance Agreement Framework (SAAF) that fuses caption‑grounded semantic responses with RGB differences for accurate change localization. Experiments on the Flair‑RSGen and WHU‑CDC datasets demonstrate that SAAF outperforms existing limited‑supervision baselines in macro‑averaged IoU and F1 metrics.
By Yuan Qian, Jie Ma
S3VD is a new video deraining framework that leverages semantic guidance and spatio‑temporal scanning to improve performance over existing State Space Models such as Mamba. It introduces a Multi‑Scale Semantic Fusion module that uses DINOv2 priors to preserve 2D spatial semantics, and a Spatio‑Temporal Scanning Fusion module that incorporates a Decoupled‑Gating Mamba layer to better model intra‑ and inter‑frame correlations. Experiments on video deraining benchmarks show that S3VD achieves state‑of‑the‑art results, improving PSNR by an average of 0.84 dB over Mamba‑based baselines.
By Kui Jiang, Yiang Chen, Yan Luo, Zhaocheng Yu, Junjun Jiang, Xianming Liu
arXiv:2606. 10328v1 Announce Type: cross Abstract: The integration of spatial and spectral information is beneficial to the improvement of change detection performance.
By Yunlong Liu, Zekai Zhang