arXiv Machine Learning By Junyi Ye, Gargi Vijay Borde

Regime-Gated Residual Mixture-of-Experts for Cross-Sectional Volatility Forecasting

Read the original on arXiv Machine Learning →

arXiv:2608. 12251v1 Announce Type: cross Abstract: Financial volatility is regime dependent, yet incorporating regime information into neural networks can also destabilize training.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Jul 13

Forking-Sequences: Statistically and Computationally Efficient Multi-Horizon Forecasting with Reduced Volatility

arXiv:2510. 04487v5 Announce Type: replace Abstract: While accuracy is a critical requirement for time series forecasting, an equally important desideratum is reasonable forecast volatility across forecast creation dates (FCDs).

By Willa Potosnak, Malcolm Wolff, Mengfei Cao, Ruijun Ma, Tatiana Konstantinova, Dmitry Efimov, Michael W. Mahoney, Boris Oreshkin, Kin G. Olivares
arXiv AI
Aug 19

MoFE: A Novel Mixture-of-Experts Framework with Fourier Neural Operators for Cryptocurrency Forecasting

MoFE is a deep learning framework that combines Fourier Neural Operators with a Mixture-of-Experts architecture to forecast cryptocurrency prices. It models volatility as a mix of multi-frequency components—including fundamental growth, mining costs, halving events, and market sentiment—using adaptive FNO and convolutional experts. Experiments on Bitcoin data from 2020 to 2025 show MoFE outperforms existing models in short‑term horizons, reducing phase‑lag errors and improving directional accuracy and information coefficient, which translates into higher Sharpe ratios in simulated trading.

By Bowen Liu, Mingming Sun
arXiv Machine Learning
Sep 18

Fast Training of Mixture-of-Experts for Time Series Forecasting via Expert Loss Integration

The paper introduces an adaptive Mixture-of-Experts (MoE) framework for time series forecasting that incorporates expert-specific losses to give each expert a direct learning signal independent of gating weights. The overall objective combines base forecasting loss with these expert losses, encouraging experts to specialize on different temporal segments. A partial online learning strategy is added for efficient incremental updates, and experiments on economic, tourism, and energy datasets show the method outperforms state‑of‑the‑art neural models and foundation models, with ablation studies confirming the benefit of expert loss integration.

By Btissame El Mahtout, Florian Ziel