Quantifying and Mitigating Self-Preference Bias of LLM Judges
DGX agentarXiv:2604.22891v1 Announce Type: cross Abstract: LLM-as-a-Judge has become a dominant approach in automated evaluation systems, playing critical roles in model alignment, leaderboard construction, qu