arXiv AI

Censoring-Aware In-Context Learning for Generalized Supplier Lead Time Estimation in Supply Chain Planning

arXiv:2607. 18530v1 Announce Type: cross Abstract: Supplier lead time forecasting is a central input to material requirements planning, inventory optimization, and supply chain risk management.

arXiv Machine Learning
Aug 27

CEDAR: Controlled and Event-Driven Demand Forecasting via Residual Decomposition

CEDAR is a two‑stage framework for demand forecasting that incorporates planned actions and external event signals. Stage I uses an Action‑Interleaved Transformer to model controllable state transitions under interventions, while Stage II applies a Residual Correction Module that aligns event descriptions with product context using LLM‑assisted text representations. Experiments on a large Alibaba 1688 dataset show that CEDAR improves simulation accuracy over traditional time‑series forecasting baselines and benefits real‑world budget planning.

By Junjie Meng, Ranxu Zhang, Zi-an Zhang, Shujun Liu, Xiaoning Qi, Xiaozhou Xu, Yanyong Zhang, Hui Xiong, Chao Wang
arXiv Machine Learning
2d ago

Beyond Model Ranking: Regime Diagnosis for Distributional-Statistical Misspecification in Industrial Time-Series Forecasting

The paper introduces a regime‑diagnosis framework for industrial time‑series forecasting, highlighting that canonical loss functions embed fixed statistical priors that are violated in real‑world demand regimes such as zero‑inflation, skewness, and high variability. It proposes the Regime‑wise Relative Bias Vector (RBV) as a metric‑agnostic diagnostic that decomposes bias into an intrinsic floor and an excess attributable to training. A large‑scale study across 13 loss objectives and 60,000+ series demonstrates that regime‑aware diagnosis distinguishes optimization‑from‑bias failures and that regime‑aware training can eliminate pooling‑induced bias that mere capacity scaling cannot.

By Pengyu Nie, Chenglang Xu, Yaoshi Chen, Chaogan Ren, Wei Hu, Chao Yang, Jiangong Zhang
arXiv Machine Learning
Sep 22

Monotone-Constrained Diffusion Models for Long-Horizon Production Forecasting

The paper introduces Physics‑SIMS‑TS, a conditional diffusion model designed for long‑horizon oil and gas production forecasting. It enforces monotone decline through negative guidance, decline‑curve constraints, and isotonic projection during sampling, and incorporates spatial training augmentation and an ensembled stochastic sampler to produce calibrated predictive distributions. Evaluated on over 35,000 wells across three jurisdictions, Physics‑SIMS‑TS achieves the highest accuracy among diffusion forecasters and matches transformer ensembles, with only a 0.5% increase in mean squared error for monotonicity.

By Temesgen Mikael Abraha, Yves Lucet
arXiv Machine Learning
5d ago

Aurora-X: Built for Extreme Time Series Forecasting

Aurora‑X is a billion‑parameter time‑series foundation model designed for extreme forecasting tasks. It employs a progressive curriculum that starts with channel‑independent pretraining, then adds cross‑variable dependencies, variable context and horizon lengths, and optional future covariates during mid‑training. A variable‑resolution post‑training stage allows adjustable temporal spans per token at inference, while a pattern‑guided mixture‑of‑experts expands capacity through sparse activation and expert specialization. An implicit quantile network head predicts arbitrary quantiles, enhancing probabilistic forecasting flexibility. Experiments on GIFT‑Eval, TIME, FEV‑Bench, TFB, and DAG‑Bench show state‑of‑the‑art performance against both pretrained TSFMs and task‑specific supervised models.

By Xingjian Wu, Chenjuan Guo, Xiangfei Qiu, Zhigang Hu, Hanyin Cheng, Peng Chen, Yang Shu, Jilin Hu, Bin Yang
arXiv Machine Learning
Aug 31

D-TAIA: Domain-Aware LLM Adaptation for Multi-Task Predictive Process Monitoring

D-TAIA is a framework that adapts large language models for multi‑task predictive process monitoring, jointly predicting the next activity and remaining time of ongoing cases. It uses domain‑aware triplet loss pre‑training, FAISS‑based nearest‑neighbor retrieval for time estimation, and a TAIA inference strategy to preserve sequential reasoning while fine‑tuning a 10 M‑parameter backbone. Across four real‑world event logs, D‑TAIA achieves state‑of‑the‑art or competitive results compared to a fine‑tuned LLM and a recurrent neural network baseline, with ablation studies showing the effectiveness of NLP and computer‑vision techniques for this domain.

By Sjoerd van Straten, Christine Jacob, Marwan Hassani