From Correctness to Utility: Gain-Based Prefix Evaluation for LLM Reasoning
DGX agentarXiv:2606.07190v1 Announce Type: new Abstract: Reasoning prefixes shape the future trajectory of LLM problem solving, yet existing process reward models usually evaluate them through local step corre