arXiv Computer Vision
2d ago

A Simulation-Grounded Agentic VLM Framework for Wildfire Monitoring and Reporting

The paper introduces a simulation‑grounded vision‑language model (VLM) framework for wildfire monitoring that converts 2D wildfire simulations into labeled video episodes using a fixed Blender mapping to create low‑detail 3D proxies. These proxies, along with controllable video generation, provide a multimodal memory that a training‑free multi‑agent VLM system uses to retrieve reference episodes, reconcile visual and memory‑based predictions, and generate structured wildfire reports. The system achieves 77.3% accuracy on six simulator‑derived report fields, outperforming direct VLM querying and text‑only memory baselines.

By Duowen Chen, Yuchen Sun, Zhiqi Li, Yuxuan Liao, Sinan Wang, Bart van Bloemen Waanders, Bo Zhu
arXiv Machine Learning
Aug 4

Obshazard-bench: Benchmarking Multimodal Foundation Models for Real-Time Disaster Intelligence from Raw Earth Observation Streams

arXiv:2608. 00012v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are increasingly used to interpret Earth observation data, yet their capability to support real-world disaster emergency response remains insufficiently evaluated.

By Fengxiang Wang, Qiuyang Yu, Yueying Li, Mingshuo Chen, Chengchi Fei, Kaiyi Xu, Lixin Gu, Wangxu Wei, Junchao Gong, Lipeng Ma, Jiong Wang, Fenghua Ling, Wenlong Zhang, Xue Yang, Wenjing Yang, Ben Fei, Long Lan
arXiv Machine Learning
Sep 17

Modular Deep Learning Mechanisms for Auditable Next-Day Wildfire Spread Prediction

The paper presents modular deep learning augmentations for next‑day wildfire spread prediction, including wind‑ and slope‑conditioned attention biases, physics‑feature retrieval‑augmented output correction, and fire‑conditioned dual‑stream gating. These modules are evaluated on five backbone models using the Next Day Wildfire Spread benchmark, with staged ablations, directional audits, retrieval perturbations, calibration measures, and computational comparisons. The best augmented SwinUNETR model achieves an F1 score of 0.4216 and an AUC‑PR of 0.3673, while a mixed ensemble reaches 0.4292 and 0.3790, demonstrating that predictive performance, operational trustworthiness, and computational practicality can be simultaneously improved.

By Miguel Esparza, Aydin Ayanzadeh Ahmad Mousavi, Ali Mostafavi
arXiv Machine Learning
Jul 24

Climate-resilient electric vehicle charging infrastructure for sustainable cities: An interpretable causal-ensemble framework for preventive maintenance and low-carbon mobility

arXiv:2607. 21444v1 Announce Type: cross Abstract: Reliable electric vehicle (EV) charging infrastructure is a cornerstone of sustainable, low-carbon cities, yet urban climate stress such as extreme heat, heavy precipitation, and humidity increasingly raises equipment fault risk and undermines the resilience of urban energy and mobility services.

By Cande Lian (School of Management, Foshan University, Foshan, China), Wentao Zeng (School of Management, Foshan University, Foshan, China), Jiabin Wu (School of Management, Foshan University, Foshan, China), Yiming Bie (School of Transportation, Jilin University, Changchun, China), Wei Zhou (Department of Civil and Environmental Engineering, National University of Singapore)
arXiv Machine Learning
Aug 20

Scalable Geospatial Machine Learning for Power-Line Asset Risk: Integrating Remote Sensing for Lightning and Vegetation Risk Modelling

The paper presents a modular, scalable framework for estimating the probability of failure (PoF) of power‑line assets using geospatial machine learning. It integrates diverse environmental predictors—topography, vegetation indices, lightning climatology, proximity features, and operational records—to model vegetation‑ and lightning‑related failure modes. The architecture is designed to be computationally efficient, easily extensible to new data sources, and suitable for large‑scale utility deployment, enabling asset‑level risk stratification for inspection and resilience planning.

By Artur Sokolovsky, Bhavik Merai, Moe Jafari, Muen Chen