arXiv:2606. 18183v1 Announce Type: cross Abstract: Temporal difference (TD) learning with linear function approximation is a core method for policy evaluation.
By M. Forzo, E. Monzio Compagnoni, A. Russo, A. Pacchiano
arXiv:2607. 27781v1 Announce Type: cross Abstract: We establish a dimension-efficient neural network approximation theory for solutions to fractional parabolic equations with lower-order drift and potential terms.
By Jae-Hwan Choi, Hyojae Lim, Jinsol Seo, Young-Jin Sim, Changhoon Song
arXiv:2607. 06935v1 Announce Type: cross Abstract: Reinforcement learning (RL) is increasingly grounded in tools from probability, optimization, and operator theory.
By Denis Belomestny, Alexander Gasnikov, Egor Gladin, Alexey Naumov, Artemy Rubtsov, Yuri Sapronov, Daniil Tiapkin, Nikita Yudin
arXiv:2412. 03405v3 Announce Type: replace-cross Abstract: Motivated by dynamic risk measures and conditional $g$-expectations, in this work we propose a numerical method to approximate the solution operator given by a Backward Stochastic Differential Equation (BSDE).
By Pere Diaz-Lozano, Giulia Di Nunno
arXiv:2410. 14788v4 Announce Type: replace-cross Abstract: Neural operator (NO) architectures learn nonlinear maps between infinite-dimensional function spaces and are widely used to accelerate simulation and enable data-driven model discovery.
By Takashi Furuya, Anastasis Kratsios
arXiv:2606. 16846v1 Announce Type: cross Abstract: We study the operator-theoretic core of Q-learning in continuous-time stochastic control with continuous states and actions.
By Qian Qi
The paper investigates how parameter Jacobians influence the stability of network outputs within the framework of network dynamics, learning models, and neural tangent kernels (NTK). It demonstrates that linearized dynamics can be expressed as a semigroup of linear operators on Hilbert spaces, and provides explicit a priori perturbation bounds for fixed‑kernel linearizations in the NTK setting. The authors also offer refinements for task‑specific spaces, ergodic comparison estimates, spectral‑distribution conditions, and extensions to nonautonomous NTK evolutions, supported by worked examples.
By Halyun Jeong, Palle E. T. Jorgensen, Hyun-Kyoung Kwon, Myung-Sin Song, James Tian
arXiv:2608.22636v1 Announce Type: cross
Abstract: Q-learning with linear function approximation can be unstable because an arbitrary approximation architecture need not preserve the Bellman contracti...
By Shengbo Wang
arXiv:2607. 14361v1 Announce Type: cross Abstract: We address fundamental challenges in representing and computing $\mathbb{R}^{d}$-valued predictable square-integrable processes over $[0,T]$, collected in the space $\mathcal{H}^2_T(\mathbb{R}^{d})$.
By Anastasis Kratsios, Giulia Livieri, Philipp Schmocker
arXiv:2606. 04275v1 Announce Type: cross Abstract: We present a novel theoretical framework for deep reinforcement learning (RL) in continuous environments by modeling the problem as a continuous-time stochastic process, drawing on insights from stochastic control.
By Saket Tiwari, Tejas Kotwal, George Konidaris
arXiv:2606. 29438v1 Announce Type: cross Abstract: In this paper, we develop a fractional stochastic neural network with residual dynamics driven by fractional Brownian motion.
By Yuecai Han, Jianming Xu
arXiv:2606. 09047v1 Announce Type: cross Abstract: A classical universal stabilization formula offers the practitioner no design freedom: it is a single, parameter-free object.
By Miroslav Krstic, Luke Bhan