arXiv Machine Learning By Donghyun Lee, Yuhang Li, Ruokai Yin, Priyadarshini Panda

KronQ: LLM Quantization via Kronecker-Factored Hessian

Read the original on arXiv Machine Learning →

arXiv:2607. 07964v1 Announce Type: new Abstract: Post-training quantization (PTQ) is a widely adopted technique for compressing large language models (LLMs) without retraining.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.