arXiv Machine Learning

Can AI Weather Models Predict Beyond Two Weeks? A Quantitative Benchmark and Analysis of Long Rollouts

arXiv:2605. 30184v2 Announce Type: replace Abstract: While AI weather models excel at short-to-medium range forecasts (up to 15 days), they frequently suffer from ill-defined "instabilities" when rolled out over longer horizons.

arXiv Machine Learning
Sep 25

Improving global precipitation forecasts with an AI weather model trained on satellite observations

The paper presents Laxmi, a retrained version of the AIFS weather model that uses satellite-based precipitation observations instead of ERA5 reanalysis data. Laxmi achieves a 19% improvement in global probabilistic accuracy, reduces drizzle overprediction by 33%, and boosts the 95th percentile Brier skill score by 57%. In a case study of 10 Indian tropical storms, Laxmi accurately forecasted 150 mm event-total precipitation in 7 events, outperforming both the original AIFS and the leading physical model IFS.

By Julian F. Schmitt, Bertrand Delorme, Robert C. King, Yashica Patodia, Tapio Schneider, Aditi Sheshadri, Ravi Jain
arXiv Machine Learning
Jun 18

Benchmarking Physics-Informed Time-Series Models for Operational Global Station Weather Forecasting

arXiv:2406. 14399v4 Announce Type: replace Abstract: The development of Time-Series Forecasting (TSF) models is often constrained by the lack of comprehensive datasets, especially in Global Station Weather Forecasting (GSWF), where existing datasets are small, temporally short, and spatially sparse.

By Tao Han, Zhibin Wen, Zhenghao Chen, Dazhao Du, Song Guo, Lei Bai
arXiv AI
Jul 7

Evaluating Skill and Stability of ArchesWeather and ArchesWeatherGen under Multi-Decadal Climate Simulations

arXiv:2605. 29976v2 Announce Type: replace-cross Abstract: We evaluate the climate simulation capabilities of ArchesWeather and ArchesWeatherGen, two machine learning models originally trained for weather forecasting and evaluated up to a 10-day lead time.

By Renu Singh, Robert Brunstein, Antonia Jost, Yana Hasson, Thomas Rackow, Claire Monteleoni, Christian Lessig, Guillaume Couairon
arXiv Machine Learning
Sep 4

Improving precipitation forecasts in an AI weather model using observational data

The paper presents a graph-transformer AI weather model that is fine‑tuned with high‑resolution IMERG precipitation observations, moving beyond the traditional reliance on the ERA5 reanalysis dataset. This approach yields up to a 19% improvement in medium‑range continuous ranked probability scores and a 57% better Brier skill score for extreme rainfall compared to leading operational models, while also excelling in tropical storm and drizzle prediction. The study demonstrates that directly incorporating observation‑based precipitation data into AI training can markedly enhance forecast accuracy, though physics‑based models still outperform for the heaviest events.

By Julian F. Schmitt, Bertrand Delorme, Robert C. King, Yashica Patodia, Tapio Schneider, Aditi Sheshadri, Ravi Jain
arXiv Statistics ML
Sep 11

Stress-Testing Dynamical and Generative Downscaling Using Subseasonal Extreme Precipitation Forecasts

The study compares the Weather Research and Forecasting (WRF) dynamical model with an unpaired diffusion-based generative model for downscaling extreme precipitation events up to three weeks ahead. Both models outperform raw European Centre for Medium-Range Weather Forecasts forecasts when evaluated against Swiss rain gauge-radar observations, but their strengths differ by atmospheric regime: WRF excels in a multicell, non‑stationary event, while the diffusion model performs more consistently and better in a stationary supercell event.

By Mauricio Lima, Marika Koukoula, Romain Pilon, Monika Feldmann, Erwan Koch, Daniela I. V. Domeisen, Tom Beucler
arXiv Machine Learning
Sep 2

GenONet: A Generative operator Network for High-Resolution Precipitation Nowcasting

GenONet introduces a Spatio-Temporal U-DeepONet architecture that serves as a generator in a GAN framework for high‑resolution precipitation nowcasting up to three hours ahead. By learning continuous‑time precipitation dynamics with a Deep Operator Network and enforcing physics through a moisture‑conservation loss, the model produces sharp, physically consistent forecasts that outperform baselines, especially for high‑intensity events and longer lead times. Ablation studies confirm the added value of the physics‑informed regularizer and the synergy of operator learning with adversarial training.

By Mohammad Kian Golkar, Luciano Alves de Oliveira, Mohammad Khanjani
arXiv Machine Learning
Sep 4

WeatherNext 3: Increasing resolution and performance of global weather models with raw observations

WeatherNext 3 is a new AI‑driven global weather model that improves both spatial and temporal resolution by generating hourly forecasts at 0.1° resolution, matching the best physics‑based models. It incorporates low‑latency geostationary satellite data and learns to predict satellite‑derived precipitation, tropical cyclones, and station observations, enabling 2 m temperature and dewpoint predictions anywhere and anytime. By directly using raw observations instead of relying solely on analysis data, WeatherNext 3 sets a new state‑of‑the‑art for probabilistic medium‑range forecasting skill.

By Stephan Rasp, Boris Babenko, Dominic Masters, Andrew El-Kadi, Samier Merchant, Guy Shalev, Ilan Price, Fred Zyda, Remi Lam, Sasha Shysheya, Matthew Willson, Stratis Markou, Shreya Agrawal, Suhani Vora, Mohammed Alewi Hassen, Sunny Mak, Tom R. Andersson, Megan Bela, Akib Uddin, Nofar Peled Levi, Ben Gaiarin, Ferran Alet, Aaron Bell, Peter Battaglia, Alvaro Sanchez-Gonzalez