arXiv Machine Learning By Jianhao Ma, Yuxin Chen

A lower bound for stepsize-based acceleration of gradient descent

Read the original on arXiv Machine Learning →

arXiv:2608. 10418v1 Announce Type: cross Abstract: Recent work has shown that, for smooth convex optimization, plain gradient descent can be accelerated from its textbook convergence rate of $O(T^{-1})$ (where $T$ denotes the number of iterations) to $O\big(T^{-\log_2(1+\sqrt{2})}\big)$ using carefully designed stepsize schedules alone, without resorting to momentum or other algorithmic modifications.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.