arXiv Machine Learning

Performance-Carbon Trade-Offs across Architectural Biases in Shear Flow Forecasting

arXiv:2509. 24517v3 Announce Type: replace Abstract: Development of modern deep learning methods has been driven primarily by the push for improving model efficacy (accuracy metrics), leading to large-scale models that require massive computational resources and result in considerable carbon footprint across the model lifecycle.

arXiv Machine Learning
Aug 20

An Empirical Benchmark of Deep Time-Series Models for Smart Meter Energy Forecasting

The paper presents an empirical benchmark of nine modern deep‑learning models for time‑series forecasting of smart‑meter energy consumption, evaluated on two publicly available datasets. It examines how historical input length, prediction horizon, and model architecture affect accuracy, finding that longer historical context improves performance up to a saturation point and that accuracy declines with longer horizons. The study also compares computational complexity, showing that lightweight architectures achieve similar performance to heavier models, and notes that model choice has limited impact across most demographic and household subgroups.

By Behnaz Kavoosighafi, Maria Eidenskog, Wiktoria Glad, Katerina Vrotsou
arXiv AI
Aug 10

Seeking SOTA: Time-Series Forecasting Must Adopt Taxonomy-Specific Evaluation to Dispel Illusory Gains

arXiv:2603. 15506v2 Announce Type: replace-cross Abstract: We argue that the current practice of evaluating AI/ML time-series forecasting models, predominantly on benchmarks characterized by strong, persistent periodicities and seasonalities, obscures real progress by overlooking the performance of efficient classical methods.

By Raeid Saqur, Christoph Bergmeir, Blanka Horvath, Daniel Schmidt, Frank Rudzicz, Terry Lyons
Hugging Face Trending Papers
Aug 19

An Empirical Benchmark of Deep Time-Series Models for Smart Meter Energy Forecasting

The paper presents an empirical benchmark of nine deep learning models for smart meter energy forecasting, evaluating them on two public datasets. It examines how historical input length, prediction horizon, and model architecture affect accuracy, finding that longer historical context improves performance up to a saturation point while accuracy declines with longer horizons. The study also compares computational cost, showing lightweight models achieve similar accuracy to heavier ones, and notes that model choice matters less across most population segments.

arXiv AI
2d ago

Beyond State-of-the-Art: Standardising Environmental Impact Metrics for AI Research

The paper highlights that as Large Language Models grow in capability and prevalence, their environmental footprint is increasing, yet the machine learning community lacks standardized carbon accounting practices. An automated review of 5,285 NeurIPS 2025 papers shows almost no reporting of environmental impact. To address this, the authors propose standardized sustainability metrics for training efficiency, heuristics for estimating inference carbon costs, a software tool called carbonbenchmark for tracking emissions, and the SMAJ framework to encourage prioritizing computational efficiency and environmental accountability over marginal accuracy gains.

By Lachlan McGinness, Dan Pagendam, Robert Offner
arXiv Machine Learning
Sep 4

SimCast-S2S: A Computationally Efficient Diffusion Model for Subseasonal Precipitation Forecasting

SimCast‑S2S is a generative latent‑diffusion framework designed for probabilistic subseasonal‑to‑seasonal precipitation forecasting. It tackles three key challenges: it uses a diffusion‑based generative pipeline for uncertainty quantification, operates in a compact latent space learned by VAEs for efficient large‑ensemble generation, and employs transfer learning with LoRA to overcome limited training data. On reanalysis data, it outperforms deep‑learning baselines and competes with or surpasses state‑of‑the‑art operational systems such as ECMWF‑S2S.

By Hiep V. Dang, Antonios Mamalakis
arXiv Machine Learning
Jun 9

TriHead-GAN: A Generative Adversarial Network with Triple-Head Discriminator for Carbon Emission Time Series Generation

arXiv:2606. 07569v1 Announce Type: new Abstract: Accurate carbon emission monitoring is critical for climate policy and emerging regulatory mechanisms such as the EU Carbon Border Adjustment Mechanism, yet city-level high-frequency monitoring data remain extremely scarce, severely limiting data-hungry deep learning models.

By Zesen Wang, Lijuan Lan, Yonggang Li, Chunhua Yang
arXiv Machine Learning
Aug 28

SimCast-S2S: An Efficient Generative Model for Subseasonal Precipitation Forecasting via Transfer Learning from Climate Simulations

SimCast‑S2S is a generative latent‑diffusion model designed for probabilistic subseasonal‑to‑seasonal precipitation forecasting. It tackles three key challenges: it uses a diffusion pipeline to capture uncertainty, operates in a compact latent space to enable efficient large‑ensemble generation, and leverages transfer learning with low‑rank adaptation to train on limited reanalysis data after pretraining on climate simulations. The model outperforms deep‑learning baselines and competes with, or surpasses, operational systems such as the ECMWF‑S2S baseline without requiring extensive post‑processing.

By Hiep V. Dang, Antonios Mamalakis