Efficient Long-Horizon Vision-Language-Action Models via Static-Dynamic Disentanglement
DGX agentarXiv:2602.03983v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for generalist robotic control. Built upon vision-language model (