GROW: Aligning GRPO with State-Action Modeling for Open-World VLM Agents
DGX agentarXiv:2605.20246v1 Announce Type: new Abstract: Recently, vision-language model (VLM) agents have shown promising progress in open-world tasks, where successful task completion often requires multiple