The Multilingual Quantization Tax: Structural Collapse and Typological Fragility in Edge SLMs
arXiv:2608.09941v1 Announce Type: new Abstract: While 4-bit weight quantization is critical for deploying Small Language Models (SLMs) on edge devices, evaluations of the resulting performance degrada