arXiv:2606. 14195v1 Announce Type: new Abstract: Kalman filters based on the Embedded Latent Transfer Operators (ELTO) emerge as novel statistical tools for sequential state estimation.
By Naichang Ke, Pongpisit Thanasutives, Yoshinobu Kawahara
arXiv:2602. 05379v2 Announce Type: replace-cross Abstract: Effective reinforcement learning (RL) for complex stochastic systems requires leveraging historical data to improve sample efficiency and accelerate policy optimization.
By Hua Zheng, Wei Xie, M. Ben Feng, Keilung Choy
arXiv:2601. 22970v2 Announce Type: replace-cross Abstract: Policies learned via continuous actor-critic methods often exhibit erratic, high-frequency oscillations, making them unsuitable for physical deployment.
By Jeong Woon Lee, Kyoleen Kwak, Daeho Kim, Hyoseok Hwang
arXiv:2607. 20521v1 Announce Type: new Abstract: The state of a dynamic system evolves over time, switching among several latent modes that govern its observable behavior.
By Lei Cao, Sihang Feng, Jixin Yan, Tao Sun, Naichen Shi
arXiv:2606. 02767v1 Announce Type: cross Abstract: Kalman filtering performance is highly sensitive to model mismatch and noise covariance tuning.
By Jiho Lee, Nisar R. Ahmed, Rebecca Russell
arXiv:2511. 23310v3 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective paradigm for post-training large language models, yet the design of its baselines and learning-rate schedules remains largely heuristic.
By Zixun Huang, Jiayi Sheng, Zeyu Zheng