Grounded Semantic Re-Binding for Robust Instruction Generalization in Vision-Language-Action Models
DGX agentarXiv:2608.02497v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models excel in robotic manipulation but suffer catastrophic performance drops when canonical instructions are simply parap