REAL: Regression-Aware Reinforcement Learning for LLM-as-a-Judge
arXiv:2603.17145v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as automated evaluators that assign numeric scores to model outputs, a paradigm known a