arXiv Machine Learning

Online Reinforcement Learning in the Met Office Unified Model through Distributed Model-Agent Coupling

The study couples the Met Office Unified Model with distributed reinforcement learning agents, using a DDPG actor that applies bounded potential‑temperature corrections across 70 vertical levels. Training is performed on ten nudged forecasts, after which the frozen policy is evaluated in a non‑nudged forecast, demonstrating numerical stability. The learned policy reduces Z₅₀₀ MAE in four of six latitude bands—up to 45.8% in the northern tropics—and decreases MSLP error by up to 27.3% in certain bands, indicating promising bias‑correction potential.

Hugging Face Trending Papers
Sep 2

Online Reinforcement Learning in the Met Office Unified Model through Distributed Model-Agent Coupling

The study couples the Met Office Unified Model with distributed reinforcement learning agents that apply bounded temperature corrections across 70 vertical levels. Training on ten nudged forecasts and evaluating on a non‑nudged run, the learned policy remains numerically stable and improves forecast accuracy, reducing Z₅₀₀ MAE by up to 45.8% in tropical bands and MSLP error by up to 27.3% in certain latitudes. This experiment demonstrates the feasibility of online RL for bias correction in operational weather models.

arXiv Machine Learning
Sep 7

Advancing Subseasonal Forecasting with Machine Learning

The paper introduces Probabilistic Bias Correction (PBC), a machine learning framework that learns to correct historical probabilistic forecasts, thereby reducing systematic errors in subseasonal weather predictions. Applied to leading dynamical and AI models from ECMWF, PBC doubles the AI system’s modest subseasonal skill and improves the operationally-debiased dynamical model for most pressure, temperature, and precipitation targets. In ECMWF’s 2025 real‑time forecasting competition, PBC’s global forecasts ranked first across all weather variables and lead times, outperforming multiple operational and ensemble models.

By Hannah Guan, Soukayna Mouatadid, Paulo Orenstein, Judah Cohen, Haiyu Dong, Zekun Ni, Jeremy Berman, Genevieve Flaspohler, Alex Lu, Jakob Schloer, Joshua Talib, Jonathan A. Weyn, Lester Mackey
arXiv Machine Learning
Jul 13

Enhancing AI and Dynamical Subseasonal Forecasts with Probabilistic Bias Correction

arXiv:2604. 16238v2 Announce Type: replace Abstract: Decision-makers rely on weather forecasts to plant crops, manage wildfires, allocate water and energy, and prepare for weather extremes.

By Hannah Guan, Soukayna Mouatadid, Paulo Orenstein, Judah Cohen, Haiyu Dong, Zekun Ni, Jeremy Berman, Genevieve Flaspohler, Alex Lu, Jakob Schloer, Joshua Talib, Jonathan A. Weyn, Lester Mackey
arXiv AI
Jul 28

AIFL: A Global Daily Streamflow Forecasting Model Using a Deterministic LSTM Pre-trained on ERA5-Land and Fine-tuned on IFS

arXiv:2602. 16579v2 Announce Type: replace-cross Abstract: Reliable global streamflow forecasting is essential for flood preparedness and water resource management, yet data-driven models often suffer from a performance gap when transitioning from historical reanalysis to operational forecast products.

By Maria Luisa Taccari, Kenza Tazi, Ois\'in M. Morrison, Andreas Grafberger, Juan Colonese, Corentin Carton de Wiart, Christel Prudhomme, Cinzia Mazzetti, Matthew Chantry, Florian Pappenberger
arXiv Machine Learning
Sep 21

Learning to Advect: A Neural Semi-Lagrangian Architecture for Weather Forecasting

arXiv:2601.21151v3 Announce Type: replace Abstract: Machine-learning approaches to weather forecasting often employ a monolithic architecture in which distinct physical mechanisms, such as advection,...

By Carlos A. Pereira, St\'ephane Gaudreault, Valentin Dallerit, Christopher Subich, Shoyon Panday, Siqi Wei, Sasa Zhang, Siddharth Rout, Eldad Haber, Raymond J. Spiteri, David Millard
arXiv Machine Learning
Jun 8

Agentic World Modeling for 6G: Near-Real-Time Generative State-Space Reasoning

arXiv:2511. 02748v2 Announce Type: replace-cross Abstract: We argue that sixth-generation (6G) intelligence is not fluent token prediction but the capacity to imagine and choose -- to simulate future scenarios, weigh trade-offs, and act with calibrated uncertainty.

By Farhad Rezazadeh, Amir Ashtari Gargari, Hatim Chergui, Sandra Lagen, Merouane Debbah, Houbing Song, Lingjia Liu