arXiv Machine Learning

When Rates Are Geometric: Rate-Certificate Transfer for Contact Splittings in Optimization

arXiv:2607. 23642v1 Announce Type: cross Abstract: Discrete optimization algorithms are often analyzed through continuous-time limiting ODEs, but a convergence certificate for the ODE is not automatically one for the discrete algorithm.

arXiv Machine Learning
Jul 7

CSympNet-ID: conformal-symplectic map learning for linearly damped Hamiltonian systems

arXiv:2607. 03339v1 Announce Type: new Abstract: Learning dissipative dynamics from discrete observations is essential for reliable long-horizon prediction and physically meaningful parameter identification.

By Jiale Gong (School of Mathematics), Pengzhan Jin (National Engineering Laboratory for Big Data Analysis and Applications, Peking University, Beijing, China), Dongyang Kuang (School of Mathematics), Lu Li (School of Mathematics), Yifa Tang (State Key Laboratory of Mathematical Sciences, Academy of Mathematics and Systems Science, Chinese Academy of Sciences, Beijing, China)
arXiv Machine Learning
Jul 28

Finite-Time Analysis of the Natural Policy Gradient in Finite-Horizon Markov Decision Processes

arXiv:2607. 22982v1 Announce Type: new Abstract: Natural Policy Gradient (NPG) is a well-established Reinforcement Learning algorithm that underlies widely used methods such as Trust Region Policy Optimization and Proximal Policy Optimization, both of which have demonstrated strong empirical success.

By Asha Barua, Sajad Khodadadian