arXiv Machine Learning By Francisco Patitucci, Aryan Mokhtari

Adaptive Optimization via Momentum on Variance-Normalized Gradients

Read the original on arXiv Machine Learning →

arXiv:2602. 10204v2 Announce Type: replace Abstract: We introduce MVN-Grad (Momentum on Variance-Normalized Gradients), an Adam-style optimizer that improves stability and performance by combining two complementary ideas: variance-based normalization and momentum applied after normalization.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.