Rethinking Output Alignment For 1-bit Post-Training Quantization of Large Language Models
DGX agentarXiv:2512.21651v2 Announce Type: replace Abstract: Large Language Models (LLMs) deliver strong performance across a wide range of NLP tasks, but their massive sizes hinder deployment on resource-cons