The paper proposes treating a neural network’s layers as time steps in a state‑space model, converting Bayesian training into a smoothing problem. By propagating Gaussian moments forward and applying a Rauch–Tung–Striebel backward pass, weight posteriors are updated in closed form without gradient iterations or replay. The authors extend prior work by introducing a cross‑covariance identity that allows full‑covariance propagation through nonlinear activations, enabling more accurate online adaptation in non‑stationary classification, dynamics learning, and vision‑language‑action policy adaptation.
By Oren Wright, Haoming Jing, Qiaoan Shen, Koichiro Niinuma, Yorie Nakahira, Jos\'e M. F. Moura
arXiv:2606. 14195v1 Announce Type: new Abstract: Kalman filters based on the Embedded Latent Transfer Operators (ELTO) emerge as novel statistical tools for sequential state estimation.
By Naichang Ke, Pongpisit Thanasutives, Yoshinobu Kawahara
arXiv:2606. 01468v1 Announce Type: cross Abstract: Due to their explicit priors and ability to model uncertainty, Bayesian methods have played a major role in dynamical latent variable modeling of single-cell neural recordings.
By JR Huml, Jonathan Wenger, John P. Cunningham
arXiv:2606. 12691v1 Announce Type: cross Abstract: Auto-regressive models have emerged as powerful tools for sequential data, from language to video.
By Yahya Sattar, Sunmook Choi, Leo Maynard-Zhang, Yassir Jedra, Maryam Fazel, Sarah Dean
arXiv:2607. 20521v1 Announce Type: new Abstract: The state of a dynamic system evolves over time, switching among several latent modes that govern its observable behavior.
By Lei Cao, Sihang Feng, Jixin Yan, Tao Sun, Naichen Shi
FILT3R is a training‑free latent filtering layer for streaming 3D reconstruction that treats recurrent state updates as stochastic state estimation in token space. It maintains per‑token variance and computes a Kalman‑style gain to balance memory retention with new observations, estimating process noise online from temporal drift of candidate tokens. Experiments show that FILT3R generalizes overwrite and gating policies, shrinking gains in stable regimes and increasing them during genuine scene changes, thereby improving long‑horizon stability for depth, pose, and 3D reconstruction.
By Seonghyun Jin, Jong Chul Ye
arXiv:2511.20413v2 Announce Type: replace-cross
Abstract: \emph{Decision-focused learning} (DFL) trains predictive models to optimize downstream decisions rather than prediction accuracy alone. While...
By Zhuojun Xie, Adam Abdin, Yiping Fang
The paper studies a two‑stage learning framework that first trains an offline model using approximate nonlinear‑least‑squares estimation and then adapts it online with a meta‑LMS algorithm to handle parameter drift in nonlinear stochastic dynamical systems. It provides an upper bound on the offline generalization error that accounts for strong data correlation and distribution shift via Kullback‑Leibler divergence, and it demonstrates that the combined offline‑online approach outperforms methods that rely solely on offline or online learning. Both theoretical analysis and empirical experiments support the claimed performance gains.
By Haizheng Li, Lei Guo
arXiv:2606. 09430v1 Announce Type: cross Abstract: Online task-free continual learning (TFCL) requires intelligent agents to sequentially accumulate knowledge from an unbounded, non-stationary data stream under strict single-pass constraints and without any explicit task identifiers.
By Mingqi Yuan, Xiaoquan Sun, Shihao Luo, Jiayu Chen
The paper introduces the Belief Flow Filter (BFF), a generative filtering framework that encodes the evolving posterior distribution directly into flow matching model weights and updates them via test‑time gradient descent. By avoiding particle representations and Gaussian assumptions, BFF aligns structurally with Bayesian filtering and targets the recursive filtering operator. Empirical results on five physical systems—including chaotic dynamics, sparse observations, and a tokamak plasma estimation task—show that BFF outperforms existing methods in most benchmark metrics.
By Ruiqi Feng, Chongyi Wang, Tao Zhang, Tailin Wu
arXiv:2602. 23050v2 Announce Type: replace Abstract: Deep state-space models (DSSMs) enable temporal predictions by learning the underlying dynamics of observed sequence data.
By Alexej Klushyn, Richard Kurle, Maximilian Soelch, Botond Cseke, Patrick van der Smagt
arXiv:2507. 08922v3 Announce Type: replace-cross Abstract: Continual learning is an online paradigm where a learner continually accumulates knowledge from different tasks encountered over sequential time steps.
By Tameem Adel