arXiv Machine Learning By Christos Petridis, Zoran Obradovic, Mladen Kezunovic

Evaluating Large Language Models for Forced Outage Risk Prediction: Benefits and Comparison to Machine Learning

Read the original on arXiv Machine Learning →

This study evaluates large language models (LLMs) for predicting weather‑related forced outage risk in a distribution grid using a zero‑shot approach without labeled training data. The task is framed as binary severity classification over 3h, 6h, and 12h horizons, leveraging six years of outage records and high‑resolution weather data from central Texas. Four zero‑shot LLMs are compared to two supervised classifiers under two input settings—current weather observations and forecast data—showing that supervised models lead on macro‑F1 and precision, while newer LLMs achieve competitive scores and offer complementary strengths in reasoning and geographic scalability.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Sep 3

OutageDiT: A Generative Foundation Model for Power Outage Forecasting and Scenario Simulation

OutageDiT is a generative foundation model that produces seven‑day power‑outage trajectories at quarter‑hour resolution, trained on nationwide outage and weather data. It uses a condition encoder to process historical context and future covariates, and a shallow flow decoder to generate full trajectories, enabling point forecasting, uncertainty quantification, and conditional event simulation. The model outperforms strong baselines on forecasting benchmarks and can transfer zero‑shot to unseen regions, linking outage simulation to operational planning under uncertainty.

By Yunqin Zhu, Feng Qiu, Yao Xie
arXiv AI
Aug 25

LLM-based Agents for Forecasting and Prediction: Methods, Training, Evaluation, and Applications

arXiv:2608.23058v1 Announce Type: new Abstract: Large language models (LLMs) now support forecasting systems that combine language-based reasoning with temporal data, evidence retrieval, external too...

By Xiaogang Xu, Jiaqi Tang, Jianmin Chen, Yingying Yan, Zhenchao Tang, Xiangxin Zhou, Xiaobin Hu, Wei Wei, Jinfeng Wu, Qifeng Chen, Lu Zhou, Jiafei Wu, Zhe Liu, Jianwei Yin, Weimin Zheng
arXiv Machine Learning
6d ago

Every Fixed Metric Has a Blind Spot: A Learned Atmospheric Critic for Scoring Forecast Realism

The paper introduces a learned atmospheric critic that discriminates between real weather data and model outputs to produce a realism score. Unlike fixed metrics, the discriminator adapts to the specific failure modes of a given model, effectively detecting various synthetic corruptions in ERA5 data. Experiments show the learned critic outperforms existing metrics and reveals that realism decreases with longer forecast lead times, favoring numerical over machine‑learning models.

By Younes Elberkennou, Dmitri Demler, Thierry Meier, Luca Rispoli, Fanny Lehmann, Joel Oskarsson