arXiv AI By Priyanka Nihalchandani, Naman Srivastava, Varun Ojha, Pandarasamy Arjunan

Personalized Federated Sparse Adaptation of Time-Series Foundation Models

Read the original on arXiv AI →

arXiv:2608. 04695v1 Announce Type: cross Abstract: Federated adaptation of time-series foundation models (TSFMs) is attractive for building energy forecasting because meter data are private, distributed, and highly non-IID.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

Hugging Face Trending Papers
Aug 5

Personalized Federated Sparse Adaptation of Time-Series Foundation Models

Federated adaptation of time-series foundation models (TSFMs) is attractive for building energy forecasting because meter data are private, distributed, and highly non-IID. However, a single parameter-sharing strategy is unlikely to serve all pretrained TSFMs or building clients: fully shared adapters can suppress building-specific temporal behavior, while fully local adaptation discards cross-building transfer.

arXiv Machine Learning
Sep 18

OceanMoE: Structured Conditional Sparse Computation for Long-Horizon Multivariate Ocean Forecasting

OceanMoE is a structured conditional sparse Mixture-of-Experts framework designed for long‑horizon multivariate ocean forecasting. It fuses cross‑variable information to build target‑specific local representations and performs content‑conditioned sparse routing at each spatial location, with the number of active experts adjusted by router confidence. Experiments on ORAS5 data show that OceanMoE reduces aggregate forecasting error and maintains lower geometric‑mean normalized RMSE compared to baselines, while expert allocation varies with prediction targets and locations.

By Yishun Zhu, Jian Wang
arXiv Machine Learning
5d ago

Aurora-X: Built for Extreme Time Series Forecasting

Aurora‑X is a billion‑parameter time‑series foundation model designed for extreme forecasting tasks. It employs a progressive curriculum that starts with channel‑independent pretraining, then adds cross‑variable dependencies, variable context and horizon lengths, and optional future covariates during mid‑training. A variable‑resolution post‑training stage allows adjustable temporal spans per token at inference, while a pattern‑guided mixture‑of‑experts expands capacity through sparse activation and expert specialization. An implicit quantile network head predicts arbitrary quantiles, enhancing probabilistic forecasting flexibility. Experiments on GIFT‑Eval, TIME, FEV‑Bench, TFB, and DAG‑Bench show state‑of‑the‑art performance against both pretrained TSFMs and task‑specific supervised models.

By Xingjian Wu, Chenjuan Guo, Xiangfei Qiu, Zhigang Hu, Hanyin Cheng, Peng Chen, Yang Shu, Jilin Hu, Bin Yang