Self-Calibrating Language Models via Test-Time Discriminative Distillation
DGX agentarXiv:2604.09624v1 Announce Type: new Abstract: Large language models (LLMs) are systematically overconfident: they routinely express high certainty on questions they often answer incorrectly. Existin