arXiv Machine Learning

Expressivity and Statistical Trade-offs in Diffusion Policy Learning

arXiv:2607. 07967v1 Announce Type: cross Abstract: Diffusion-based policies have recently emerged as powerful policy parameterizations for reinforcement learning, representing state-conditioned action distributions as terminal laws of diffusion processes with parameterized drifts.

arXiv Machine Learning
Jul 22

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces

arXiv:2607. 18554v1 Announce Type: cross Abstract: We develop the Continuous Distributed Coupled Policy Gradient (CDCPG) algorithm for cooperative reinforcement learning in networked Markov decision processes with continuous state and action spaces.

By Dongming Wang, Pengcheng Dai, Wenwu Yu, Wei Ren
arXiv Machine Learning
Jul 21

Concentration and Mean-Square Bounds for Contractive Stochastic Approximation: A Unified Elementary Approach

arXiv:2607. 17595v1 Announce Type: new Abstract: We establish mean-square and concentration bounds for stochastic approximation (SA) with arbitrary norm contractive mappings, under a multiplicative noise model where the noise may scale affinely with the norm of the iterates, and the iterates are potentially unbounded.

By Siddharth Chandak
arXiv Machine Learning
Jul 30

Learning Controlled Stochastic Differential Equations

arXiv:2411. 01982v2 Announce Type: replace-cross Abstract: We study the problem of learning controlled stochastic differential equations (SDEs) \[ dX_t = b(t,X_t,u_t)\,dt + \sigma(t,X_t,u_t)\,dW_t, \] whose drift and diffusion depend nonlinearly on time, state, and control values.

By Luc Brogat-Motte, Riccardo Bonalli, Alessandro Rudi