RewardBench 2: Advancing Reward Model Evaluation
DGX agentarXiv:2506.01937v2 Announce Type: replace Abstract: Reward models are used throughout the post-training of language models to capture nuanced signals from preference data and provide a training target