Hugging Face Trending Papers

GISAgentBench: A Practitioner-Sourced Benchmark for Evaluating LLM Agents on GIS Tasks

Read the original on Hugging Face Trending Papers →

Geographic Information System (GIS) professionals rely on multi-step spatial analysis workflows to support decision-making in urban planning, disaster response, and environmental monitoring. The process is tedious, time-consuming, and error-prone.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at Hugging Face Trending Papers.

arXiv AI
Sep 15

ANASSA: An Agentic AI Orchestration Framework for Spatial Intelligence

The paper introduces ANASSA, an agentic AI orchestration framework designed for spatial intelligence in geographic information systems. It addresses gaps in current systems by integrating structured spatial reasoning, multi‑agent workflow orchestration, execution feedback, authoritative validation, provenance, uncertainty handling, and human decision authority. The architecture is detailed with eleven components across four layers, a six‑step Geospatial AI Cognitive Loop, cross‑component contracts, and governance mechanisms to ensure traceability, reproducibility, and accountability.

By Constantinos Papantoniou, Brian Hilton
arXiv AI
Jun 12

GeoNatureAgent Benchmark: Benchmarking LLM Agents for Environmental Geospatial Analysis Across Frontier and Open-Weight Foundation Models

arXiv:2606. 12821v1 Announce Type: new Abstract: Environmental scientists spend disproportionate effort on data wrangling rather than analysis, and AI agents that automate geospatial workflows remain unvalidated: no benchmark evaluates agents operating through structured tool calling against real APIs.

By Gabriel Diaz-Ireland, Diego Prieto-Herr\'aez, Mario Garc\'ia Peces, Javier Vel\'azquez, Devika Jain
arXiv AI
Jun 12

TerraBench: Can Agents Reason Over Heterogeneous Earth-System Data?

arXiv:2606. 13148v1 Announce Type: new Abstract: Climate and environmental decision-making increasingly requires reasoning across heterogeneous inputs, including gridded physical data, satellite imagery, geospatial context, and simulator outputs.

By Dat Tien Nguyen, Thao Nguyen, Fadillah Adamsyah Maani, Huy M. Le, Muhammad Umer Sheikh, Numan Saeed, Muhammad Haris Khan, Salman Khan