arXiv Machine Learning By Jiayu Zhang, Tianyi Lin

Scale-Invariant Neural Network Optimization: Norm Geometry and Heavy-Tailed Noise

Read the original on arXiv Machine Learning →

arXiv:2605. 18528v2 Announce Type: replace-cross Abstract: A growing lesson from neural network optimization is that optimizer design should respect how the model is parametrized.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.