RARM: Confidence-Gated Progress Reward Modeling for RL in Manipulation
DGX agentarXiv:2606.22027v1 Announce Type: new Abstract: Reinforcement learning for robot manipulation is often bottlenecked by reward design, especially in long-horizon tasks: sparse success rewards provide w