Q-MINO: A Minimal-Norm Method for Quantization-Aware Training
Read the original on arXiv Machine Learning →The paper introduces Q-MINO, a Quantization-Aware Minimal-Norm Optimizer designed to improve training of ultra-low-bit neural networks. Q-MINO uses a temporal bundle method that incorporates gradient consensus, state-drift regularization, and an alignment constraint to produce stabilized, minimum-norm update directions. The authors solve the resulting constrained subproblem with a warm-started Frank–Wolfe procedure and provide theoretical convergence guarantees via a stochastic Lyapunov Kurdyka–Łojasiewicz framework, along with numerical experiments demonstrating its effectiveness across various quantization levels.
Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.