arXiv Computation and Language

LLM-Driven Training-free Location-Attribute Synergic Fusion: A Closed-Loop Paradigm for Dual-source Encrypted POIs and LULC Mapping

The paper introduces a novel, training‑free, LLM‑driven closed‑loop framework for fusing dual‑source encrypted points of interest (DSEP) to improve land‑use/land‑cover (LULC) mapping. By iteratively refining location transformations and attribute correspondences—using LLM‑based matching, particle swarm optimization, and fuzzy attribute reassessment—the method reduces matching complexity from O(N²) to O(N) and converges in about two iterations. Experiments across 31 Chinese provincial capitals demonstrate a 4.58 m average location residual and 95.12 % attribute accuracy, outperforming existing baselines by 1.77 m and 14.87 % respectively, and enabling georeferencing of encrypted vector data to WGS‑84 without ground control points.

arXiv Computer Vision
Sep 25

An Automated Georeferencing Technique for Multi-Temporal Stope Point Clouds for Downstream Geotechnical Analysis

The paper introduces 3D-TARGeT, an automated method that uses inexpensive rectangular tags to register and georeference multi‑temporal UAV laser scans of underground mine stopes. By combining tag detection, geometric matching, and rigid transformation estimation, the technique achieves centimetre‑level accuracy, outperforming conventional automatic registration methods. The authors validated the approach on four simulated stope scans, demonstrating its potential to streamline data integration for geotechnical analysis and mine planning.

By Dibyayan Patra, Simit Raval, Pasindu Ranasinghe, Bikram Banerjee, Ismet Canbulat
arXiv Computer Vision
Sep 1

Multi-Sensor Mapping of Vulnerable Urban Settlements Using SAR, Multispectral, and Hyperspectral Imagery: A Case Study in C\'ordoba, Argentina

The paper introduces a multi‑sensor deep learning framework for mapping informal settlements in Córdoba, Argentina, using high‑resolution PlanetScope multispectral imagery, COSMO‑SkyMed SAR data, and medium‑resolution PRISMA hyperspectral observations. It compares SAR‑only, MS‑only, and various fusion strategies (early, middle, late) and finds that late fusion with hyperspectral data (LF+HS) delivers the best balance of classification accuracy and spatial precision. The study also demonstrates that detections outside official polygons align with broader municipal vulnerability layers and that identified settlements show higher surface temperatures during a heatwave, highlighting localized heat amplification.

By Luigi Russo, Anabella Ferral, Silvia Liberata Ullo, Paolo Gamba
arXiv AI
Sep 17

A Systematic Evaluation of the COTQ Provincial Land Cover Product: Structural Consistency, Spectral Separability, and Relative Positioning Against ESA, ESRI, and Google Products

The paper evaluates the Quebec-specific 10‑m land‑cover product COTQ against three global 10‑m datasets (ESA WorldCover, ESRI LandCover, and Google DynamicWorld). Using structural indicators, spectral separability metrics, and photo‑interpretation, the study finds that COTQ most closely resembles ESA WorldCover but shows systematic differences in urban, wetland, and rocky classes. The analysis clarifies COTQ’s relative strengths and weaknesses for operational land monitoring in Quebec.

By \'Etienne Clabaut, Samuel Foucher, Yacine Bouroubi
arXiv Computer Vision
Aug 27

Synergising Local Geo-Environmental Characteristics with Spatial Context for Enhancing Landslide Susceptibility Mapping

The paper introduces a Local-Geo and Spatial Context Fusion (LGSCF) strategy that combines point-based geo-environmental features with surrounding spatial context using a feature-wise modulation mechanism. Applied to nine CNN models over a 2644 km² area in Taiwan, LGSCF consistently outperforms baseline models, achieving F1-scores up to 87.09% and AUC values up to 0.9472. The resulting susceptibility maps more accurately concentrate known landslides in high-risk zones with fewer misclassifications.

By Yusen Cheng, Lei Fan, Qinfeng Zhu, Cheng Zhang, Yangyang Li, Ron Mahabir
arXiv Machine Learning
Sep 16

From Foundation Embeddings to Cropland Maps: Label Efficiency, Temporal Transferability and Independent Human Validation

The study evaluates the use of frozen geospatial foundation embeddings (AlphaEarth) for mapping cultivated versus non‑cultivated land in Maine. Using 192 spatially separated patches and USDA Cropland Data Layer labels, a lightweight classifier achieved 93.7% overall accuracy without fine‑tuning, and a nearest‑class‑centroid rule reached 90.2%. A balanced sample of 60,000 labeled pixels was nearly as effective as the full 8.6 million‑pixel pool, and classifiers trained in one year remained accurate across 2018‑2023. In a blind human validation of 385 points, the AlphaEarth‑plus‑random‑forest map matched 95.3% of the consensus, outperforming the CDL reference (91.7%).

By Mohammad Ammar Mughees, Giovanni Montefoschi, Zhongxin Chen, Maria Antonia Brovelli
arXiv AI
Jun 15

Fusion of Pervasive RF Data with Spatial Images via Vision Transformers for Enhanced Mapping in Smart Cities

arXiv:2508. 03736v2 Announce Type: replace-cross Abstract: In this paper, we present a deep learning-based approach that integrates the DINOv2 architecture to improve building mapping by combining (possibly erroneous) maps from open-source platforms with pervasive radio frequency (RF) data collected from multiple wireless user equipments and base stations.

By Rafayel Mkrtchyan, Armen Manukyan, Hrant Khachatrian, Theofanis P. Raptis
arXiv Computer Vision
Sep 25

OptiSAR-Net++: A Large-Scale Benchmark and Transformer-Free Framework for Cross-Domain Remote Sensing Visual Grounding

OptiSAR-Net++ introduces a new cross‑domain remote sensing visual grounding task (CD‑RSVG) and the first large‑scale benchmark dataset, OptSAR‑RSVG. The framework replaces Transformer decoding with a CLIP‑based contrastive approach, employing a patch‑level Low‑Rank Adaptation Mixture of Experts for efficient cross‑domain feature decoupling and a text‑guided dual‑gate fusion module for improved semantic‑visual alignment. Experiments show state‑of‑the‑art performance on OptSAR‑RSVG and DIOR‑RSVG, with notable gains in localization accuracy and computational efficiency.

By Xiaoyu Tang, Jun Dong, Jintao Cheng, Rui Fan
arXiv AI
Jun 16

GeoRoPE: Ground-Aware Rotary Adaptation for Remote Sensing Foundation Models

arXiv:2606. 14760v1 Announce Type: cross Abstract: Remote-sensing foundation models (RSFMs) benefit from pretraining on imagery from multiple sensors and ground sampling distances (GSDs), but such exposure alone does not resolve scale mismatch during downstream adaptation.

By Yu Luo, Kun Hu, Mengwei He, Xiaogang Zhu, Shan Zeng, Allen Benter, Wei Xiang, Patrick Filippi, Thomas Francis Bishop, Zhiyong Wang
arXiv Computer Vision
Sep 24

A comparative assessment of global building and settlement datasets across geographic and settlement contexts

arXiv:2609.28154v1 Announce Type: new Abstract: Global building and settlement datasets increasingly support population mapping, exposure assessment, urban monitoring, and other analyses of the built...

By Rufai Omowunmi Balogun, Caroline Margaux Gevaert, Capucine Riom, Derrick Mirindi, Aaron Opdyke, Hamed Alemohammad, Pierre Chrzanowski, Edward Charles Anderson