Attention-Guided Layer Selection for Contrastive Decoding in Large Language Models
DGX agentarXiv:2607.23067v1 Announce Type: cross Abstract: Contrastive decoding methods such as DoLa improve the factuality of Large Language Models (LLMs) by contrasting the output distributions of mature and