BOKBO (Best of K Bad Options): Calibrated Abstention for VLA Policies
DGX agentarXiv:2605.30660v1 Announce Type: new Abstract: Test-time scaling for vision-language-action (VLA) policies, methods such as RoboMonkey, SEAL, MG-Select, and V-GPS, samples K candidate action chunks a