Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation
arXiv:2606.24633v1 Announce Type: new Abstract: Human demonstrations for robot imitation learning often contain mistakes and corrective behaviors, such as imprecise grasps, object misalignment, unstab