DriveMA: Rethinking Language Interfaces in Driving VLAs with One-Step Meta-Actions
arXiv:2605.21273v1 Announce Type: new Abstract: Driving Vision-Language-Action Models (Driving VLAs) commonly introduce natural-language reasoning as an intermediate interface for end-to-end planning,