arXiv AI

Distilled Roads: Generalisable Road Network Extraction Across Sensors, Resolutions, and Region

arXiv:2608. 03407v1 Announce Type: cross Abstract: Road network segmentation from satellite imagery remains challenging due to large geographic variation in road appearance, occlusions, and domain shifts introduced by differing resolutions and sensors.

arXiv AI
Sep 2

Restrict, Don't Retrain: Inference-Time VLM Guidance for Zero-Shot Aerial Segmentation

The paper proposes a method called Restrict, Don't Retrain that enhances zero-shot aerial segmentation by using inference-time guidance from a vision‑language model (VLM). It combines a frozen foundation model that labels every pixel with two VLM queries: one to select relevant classes and another to locate small objects missed by the base model. Experiments on four aerial datasets show consistent performance gains at each stage where the base model is competent.

By Teresa DiMeola, Charles Walter, Hong Xiao
Hugging Face Trending Papers
Aug 11

GeoSeg-OV: Bridging Geospatial Gaps with Structural Guidance for Open-Vocabulary Remote Sensing Segmentation

Open-vocabulary remote sensing segmentation has recently emerged as a promising paradigm that enables pixel-level recognition of arbitrary categories specified by natural language, including classes unseen during training. However, geospatial domain shifts caused by heterogeneous regions, spatial resolutions, and acquisition platforms weaken visual-text matching and limit cross-dataset generalization.

arXiv Computer Vision
Sep 15

Global-Local Contextual Progressive Expansion Network for Martian Landslide Segmentation in Multimodal Remote Sensing Imagery

arXiv:2609.13332v1 Announce Type: new Abstract: Automated landslide segmentation on Mars is one of the important tasks for understanding its surface processes, and all will aid in future space explor...

By Leo Thomas Ramos, Sidike Paheding, Abel A. Reyes-Angulo, Rajaneesh A., Sajinkumar K. S., Angel D. Sappa, Thomas Oommen
arXiv Computer Vision
Aug 27

OpenCVL: An Open, Diverse, and Large-Scale Dataset for Fine-Grained Cross-View Localization

OpenCVL is a large, open dataset for fine-grained cross-view localization, comprising 617,388 ground‑aerial image pairs from 41 European cities. It blends high‑end sensor data with diverse in‑the‑wild images and includes a curation framework to correct pose annotations, enabling reliable evaluation. The dataset also offers cross‑area and snowy test sets to probe generalization, and experiments show that adding noisy in‑the‑wild data improves model performance on clean tests.

By Zimin Xia, Mubariz Zaffar, Junsheng Fu, Alexandre Alahi, Julian F. P. Kooij
arXiv Machine Learning
Sep 23

MIND the Gap: A Geographic Implicit Neural Representation with Adjustable Spatial Scale

The paper introduces MIND, a method that distills specialist geospatial model embeddings into a single generalist coordinate embedding with adjustable spatial granularity, using nested supervision across multiple embedding dimensions. MIND’s design allows downstream predictors to use only leading chunks or apply a Chunked Penalty to downweight finer details without retraining the INR. The authors evaluate MIND on CoordBench, a large INR benchmark of 52 datasets and 78 targets, and report that MIND and its Chunked Penalty variant achieve the highest regression and classification scores, especially under regional holdout, establishing a new state‑of‑the‑art for geographic implicit neural representations.

By Isaac Corley, Arjun Rao, Esther Rolf, Konstantin Klemmer, Evan Shelhamer, Nils Lehmann, Marc Ru{\ss}wurm, Gengchen Mai, Nathan Jacobs, Hannah Kerner
arXiv AI
Jun 2

Beyond Visual Fidelity: Benchmarking Super-Resolution Models for Large-Scale Remote Sensing Imagery via Downstream Task Integration

arXiv:2605. 00310v2 Announce Type: replace-cross Abstract: Super-resolution (SR) techniques have made major advances in reconstructing high-resolution images from low-resolution inputs.

By Zhili Li, Kangyang Chai, Zhihao Wang, Xiaowei Jia, Yanhua Li, Gengchen Mai, Sergii Skakun, Dinesh Manocha, Yiqun Xie