arXiv Machine Learning By Viet Vu, Renyuan Xu, Jiacheng Zhang, Yufei Zhang

Expressivity and Statistical Trade-offs in Diffusion Policy Learning

Read the original on arXiv Machine Learning →

arXiv:2607. 07967v1 Announce Type: cross Abstract: Diffusion-based policies have recently emerged as powerful policy parameterizations for reinforcement learning, representing state-conditioned action distributions as terminal laws of diffusion processes with parameterized drifts.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Jul 22

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces

arXiv:2607. 18554v1 Announce Type: cross Abstract: We develop the Continuous Distributed Coupled Policy Gradient (CDCPG) algorithm for cooperative reinforcement learning in networked Markov decision processes with continuous state and action spaces.

By Dongming Wang, Pengcheng Dai, Wenwu Yu, Wei Ren
arXiv Machine Learning
Jul 21

Concentration and Mean-Square Bounds for Contractive Stochastic Approximation: A Unified Elementary Approach

arXiv:2607. 17595v1 Announce Type: new Abstract: We establish mean-square and concentration bounds for stochastic approximation (SA) with arbitrary norm contractive mappings, under a multiplicative noise model where the noise may scale affinely with the norm of the iterates, and the iterates are potentially unbounded.

By Siddharth Chandak