MixQuant: Adaptive Mixed-Precision Quantization for Large Language Models
arXiv:2607.23047v1 Announce Type: cross Abstract: Mixed-precision quantization improves the accuracy of post-training quantization by allocating higher bitwidths to sensitive layers, but existing meth