arXiv AI

Optimal Post-Training Quantization Scales and Where to Find Them

arXiv:2606. 10890v1 Announce Type: cross Abstract: Post-training quantization (PTQ) compresses large language models by mapping weights to low-bit representations.