arXiv Machine Learning By Miguel P. Bento, Jo\~ao Seabra

GaugeQuant: Online Learning of Quantization-Optimal Bases from LLM Symmetries

Read the original on arXiv Machine Learning →

arXiv:2607. 20757v1 Announce Type: new Abstract: Transformers are known to have internal continuous symmetries that leave outputs invariant, while modifying quantization.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.