arXiv Machine Learning By Maria Matveev, Vit Fojtik, Hung-Hsu Chou, Gitta Kutyniok, Johannes Maly

Conflicting Biases at the Edge of Stability: Norm versus Sharpness Regularization

Read the original on arXiv Machine Learning →

arXiv:2505. 21423v3 Announce Type: replace Abstract: The remarkable generalization properties of overparameterized networks are often attributed to implicit biases, such as norm minimization at small learning rates and low sharpness in the Edge-of-Stability regime.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.