Flatness Preserves Instruction Following in Vision-Language-Action Models
DGX agentarXiv:2606.23641v1 Announce Type: new Abstract: Vision-language-action (VLA) models have the potential for open-world generalization by leveraging pretrained vision-language representations, yet downs