arXiv Machine Learning By Steven Heilman, Sampad Mohanty

On the Convergence of Adam, Revisited

Read the original on arXiv Machine Learning →

arXiv:2607. 03519v1 Announce Type: new Abstract: We show that projected Adam for online optimization with arbitrary moment decay parameters $\beta_1,\beta_2\in[0,1)$ can have average regret bounded away from zero.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Sep 2

Uniform a priori bounds and error analysis for the Adam stochastic gradient descent optimization method

The paper establishes uniform a priori bounds for the Adam optimizer, enabling an unconditional error analysis for a broad class of strongly convex stochastic optimization problems. Prior analyses were conditional, assuming Adam remained bounded, whereas this work removes that assumption. The results provide a rigorous foundation for Adam’s performance in training deep neural networks and other convex optimization tasks.

By Steffen Dereich, Thang Do, Arnulf Jentzen