GUARD: Grounding Uncertainty and Ablation-Based Risk Detection for Diffusion-Based VLAs
DGX agentarXiv:2608.04510v1 Announce Type: cross Abstract: Diffusion-based vision-language-action (VLA) policies can generate plausible actions even when their predictions are weakly grounded in the visual and