arXiv Machine Learning

Interventional Flow Matching: Prospective Dose-Response Forecasting with Velocity-Field Jacobian Regularization

arXiv:2606. 29386v1 Announce Type: new Abstract: Predicting a patient's physiological trajectory under a planned treatment sequence is a prospective interventional problem, not standard time-series extrapolation.

arXiv Machine Learning
Sep 11

Evaluating Time-Series Foundation Models and Multimodal Dietary Context for CGM Forecasting

The study evaluates time‑series foundation models for continuous glucose monitoring (CGM) forecasting across eight public datasets covering Type 1, Type 2, and non‑diabetes populations. Zero‑shot foundation models did not consistently beat strong task‑specific baselines, but lightweight fine‑tuning of models like Chronos‑Bolt improved root‑mean‑square error by up to 18% in both in‑distribution and out‑of‑distribution settings. Incorporating multimodal dietary context via CGMacros and a residual‑based fusion framework further reduced overall RMSE by ~3% and postprandial RMSE by ~15%, indicating that dietary signals add clinically meaningful value beyond CGM alone.

By Bowen Zhang, Hsiu-Wen Cheng, Hongyu Yang, Evie L. Shen, Joleen Vansomphone, Yuna Li, Kerry Zhou, Zitian Qu, Suning Zhao, Xiangning Deng, Hua Zhou, Jin J. Zhou
arXiv AI
1d ago

On the Divergence of Accuracy and Mechanism Consistency in Time Series World Models

The paper introduces a formal framework and benchmark for time‑series world models (TSWMs) that separates state, actions, and exogenous inputs, and defines a new metric called mechanism consistency to evaluate whether model predictions move in the expected direction when actions change. Experiments on eight public datasets show that using a frozen latent prediction space and gated output fusion improves prediction accuracy, while prediction error and mechanism consistency often diverge, with the best‑performing models sometimes failing to exhibit consistent directional responses. Adding a directional supervision loss significantly boosts mechanism consistency without affecting mean‑absolute error, providing a practical recipe for building more reliable TSWMs.

By Haochen Zhang, Jiaheng Guo, Zhen Xu, Zachary Plotkin, Nicholas Konz, Zhen Tan, Tianlong Chen
arXiv Machine Learning
Jun 2

MedGym:A Unified Continuous-Time Benchmark for Dynamic Medical Treatment Reinforcement Learning

arXiv:2606. 01028v1 Announce Type: new Abstract: Medical treatment recommendation poses several challenges to reinforcement learning (RL): patient physiology evolves in continuous time, measurements and interventions are performed at irregular intervals, and treatment effects vary substantially across individuals.

By Yuepeng Wang, Ken Kawano, Yongqi Zhou, Yoshihiko Fujisawa, Richard Weiss, Akifumi Wachi, Katsuki Fujisawa, Ying Chen, Mehrshad Sadria, Xin Liu, Kyoung-Sook Kim, Xiao Hu, Sebastien Gros, Xun Shen