arxivcs.AI2026-07-09
The Illusion of Equivalency: Statistical Characterization of Quantization Effects in LLMs
Baha Rababah, Cuneyt Gurcan Akcora, Carson K. Leung
Post-training quantization is widely used to deploy large language models in resource-constrained settings, yet its evaluation relies almost exclusively on accuracy and perplexity. We show that these metrics fail to capture behavioral changes induced by quantization. We introduce…