arXiv Machine Learning By Marcel K\"uhn, Bernd Rosenow

Explaining Near-Zero Hessian Eigenvalues Through Approximate Symmetries in Neural Networks

Read the original on arXiv Machine Learning →

arXiv:2607. 07845v1 Announce Type: new Abstract: The Hessian of the training loss governs the local geometry of the loss landscape, yet despite existing explanations for its largest eigenvalues, the origin of the vast multitude of vanishingly small eigenvalues remains elusive.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.