arXiv Machine Learning By Zijian Liu

Random Reshuffling Dominates Stochastic Gradient Descent

Read the original on arXiv Machine Learning →

arXiv:2606. 32005v1 Announce Type: cross Abstract: Stochastic Gradient Descent ($\textsf{SGD}$) is one of the most classical optimization algorithms with favorable theoretical guarantees, yet the practical implementation of $\textsf{SGD}$ differs subtly from its well-known form and is often referred to as Shuffling Stochastic Gradient Descent ($\textsf{Shuffling SGD}$).

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.