The paper presents modular deep learning augmentations for next‑day wildfire spread prediction, including wind‑ and slope‑conditioned attention biases, physics‑feature retrieval‑augmented output correction, and fire‑conditioned dual‑stream gating. These modules are evaluated on five backbone models using the Next Day Wildfire Spread benchmark, with staged ablations, directional audits, retrieval perturbations, calibration measures, and computational comparisons. The best augmented SwinUNETR model achieves an F1 score of 0.4216 and an AUC‑PR of 0.3673, while a mixed ensemble reaches 0.4292 and 0.3790, demonstrating that predictive performance, operational trustworthiness, and computational practicality can be simultaneously improved.
By Miguel Esparza, Aydin Ayanzadeh Ahmad Mousavi, Ali Mostafavi
arXiv:2606. 04143v1 Announce Type: cross Abstract: Accurate flood forecasting is essential for mitigating disaster risks and protecting communities.
By Tewodros Syum Gebre, Jagrati Talreja, Leila Hashemi-Beni
The paper introduces Probabilistic Bias Correction (PBC), a machine learning framework that learns to correct historical probabilistic forecasts, thereby reducing systematic errors in subseasonal weather predictions. Applied to leading dynamical and AI models from ECMWF, PBC doubles the AI system’s modest subseasonal skill and improves the operationally-debiased dynamical model for most pressure, temperature, and precipitation targets. In ECMWF’s 2025 real‑time forecasting competition, PBC’s global forecasts ranked first across all weather variables and lead times, outperforming multiple operational and ensemble models.
By Hannah Guan, Soukayna Mouatadid, Paulo Orenstein, Judah Cohen, Haiyu Dong, Zekun Ni, Jeremy Berman, Genevieve Flaspohler, Alex Lu, Jakob Schloer, Joshua Talib, Jonathan A. Weyn, Lester Mackey
arXiv:2608. 01864v1 Announce Type: cross Abstract: Predicting drought risk is essential for anticipating impacts on water resources, agriculture, ecosystems, and climate adaptation planning.
By Henri Funk, Cornelia Gruber, G\"oran Kauermann, Helmut K\"uchenhoff, Magdalena Mittermeier
arXiv:2607. 21597v2 Announce Type: replace Abstract: Evaluating wildfire risk systems using standard machine-learning metrics such as F1-score or IoU is fundamentally flawed: these metrics assess event prediction accuracy, not the operational coherence of a continuous risk signal.
By Nicolas Caron, Christophe Guyeux, Hassan Noura, Maxime Coulmeau, Benjamin Aynes
arXiv:2608. 05265v1 Announce Type: new Abstract: Prediction of post-wildfire debris flows is critical for mitigating hazards to communities, infrastructure, and resources during intense rainfall in recently burned areas.
By Quinn Ledingham, Zhengsen Xu, Yimin Zhu, Zack Dewis, Mabel Heffring, Saeid Taleghanidoozdoozan, Motasem Alkayid, Megan Greenwood, Lincoln Linlin Xu
arXiv:2604. 16238v2 Announce Type: replace Abstract: Decision-makers rely on weather forecasts to plant crops, manage wildfires, allocate water and energy, and prepare for weather extremes.
By Hannah Guan, Soukayna Mouatadid, Paulo Orenstein, Judah Cohen, Haiyu Dong, Zekun Ni, Jeremy Berman, Genevieve Flaspohler, Alex Lu, Jakob Schloer, Joshua Talib, Jonathan A. Weyn, Lester Mackey
WildfireSpreadBench evaluates machine‑learning models for predicting next‑day wildfire spread, comparing five discriminative and one generative architecture on the WildfireSpreadTS dataset. The study shows that model rankings differ markedly when using Average Precision versus threshold‑dependent metrics such as F1 and IoU, revealing three distinct prediction profiles—over‑predicting, balanced, and under‑predicting—that AP alone cannot distinguish. Expanding input channels modestly affected AP, underscoring that AP may favor models with predictions poorly suited for operational use.
By Arin Gopakumar, Marco Pannozzo
arXiv:2607. 21080v1 Announce Type: new Abstract: Long-horizon weather forecasting is a fundamental challenge in atmospheric science, for which autoregressive Deep Learning Weather Prediction (DLWP) has emerged as the primary paradigm.
By Yun-Ye Cai, Hsuan-Tien Lin
arXiv:2603. 11229v2 Announce Type: replace-cross Abstract: Machine learning forecast systems are moving beyond point predictions to full predictive distributions for future outcomes y conditional on complex inputs x.
By Elizabeth Cucuzzella, Rafael Izbicki, Ann B. Lee
The study presents a deployment‑aware framework for forecasting spring discharge and groundwater levels in the Edwards Aquifer over 1‑12 week horizons using 79 years of hydroclimatic data. Five machine‑learning families—extreme gradient boosting, extremely randomized trees, LSTM, CNN, and Transformers—were compared, with extreme gradient boosting consistently delivering the highest reliability (R² ≥ 0.94) and strong agreement with operational drought thresholds. The validated models were integrated into a five‑agent operational architecture that automates data acquisition, model selection, prediction, threshold monitoring, verification, literature retrieval, and reporting.
By Pramod Lekhak, Chetan Sharma, Hakan Ba\c{s}a\u{g}ao\u{g}lu, F. Paul Bertetti, Debaditya Chakraborty
arXiv:2607. 07951v1 Announce Type: new Abstract: Wildfire smoke events produce extreme PM$_{2.
By Yongcan Huang, Li Jiang, Ze Yu Liu