arXiv AI By Chenxi Zhou, Pengfei Cao, Jinyu Ye, Bohan Yu, Haida Yu, Jiang Li, Jun Zhao, Kang Liu

Quantization Degradation in Large Language Models: A Signal-Noise Perspective

Read the original on arXiv AI →

arXiv:2608. 08188v1 Announce Type: new Abstract: Post-training quantization reduces the deployment cost of large language models, yet how severely a quantized model degrades is not determined by bit-width alone.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.