arXiv Machine Learning By Mohammad Ammar Mughees, Giovanni Montefoschi, Zhongxin Chen, Maria Antonia Brovelli

From Foundation Embeddings to Cropland Maps: Label Efficiency, Temporal Transferability and Independent Human Validation

Read the original on arXiv Machine Learning →

The study evaluates the use of frozen geospatial foundation embeddings (AlphaEarth) for mapping cultivated versus non‑cultivated land in Maine. Using 192 spatially separated patches and USDA Cropland Data Layer labels, a lightweight classifier achieved 93.7% overall accuracy without fine‑tuning, and a nearest‑class‑centroid rule reached 90.2%. A balanced sample of 60,000 labeled pixels was nearly as effective as the full 8.6 million‑pixel pool, and classifiers trained in one year remained accurate across 2018‑2023. In a blind human validation of 385 points, the AlphaEarth‑plus‑random‑forest map matched 95.3% of the consensus, outperforming the CDL reference (91.7%).

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Sep 24

Geospatial embeddings detect old-growth forests but buffered spatial validation narrows their advantage over Sentinel features

arXiv:2609.28194v1 Announce Type: new Abstract: Old-growth forests develop over centuries under minimal anthropogenic disturbance, producing structurally complex and biodiverse stands. In Europe, pro...

By Thomas Ratsakatika (Department of Geography, University of Cambridge, Cambridge, UK), Mihai Zotta (Fundatia Conservation Carpathia, Brasov, Romania), Srinivasan Keshav (Department of Computer Science and Technology, University of Cambridge, Cambridge, UK), Emily R. Lines (Department of Geography, University of Cambridge, Cambridge, UK)
arXiv AI
Sep 1

A Composition-Aware Pretraining Framework for Geospatial Foundation Models

The paper introduces a composition‑aware pretraining framework for geospatial foundation models that explicitly encodes fractional land‑cover mixtures as histogram targets for each satellite image cell. By using Earth Mover’s Distance to distill these composition targets into a 36.8 M‑parameter backbone, the authors demonstrate significant improvements on region‑level tasks such as zero‑shot image retrieval and scene classification, while maintaining competitive performance on fine‑grained tasks like segmentation and object detection. The method outperforms larger models (SatMAE and Prithvi‑EO‑2.0) and achieves a 55.6 % relative boost on the ForestNet‑12 dataset, evidencing the benefit of explicit composition modeling.

By Aryan Kashyap Naveen, Abhishek Srinivas, Pranav Moothedath, Shrutilipi Bhattacharjee