CAC-VLA: Context-Gated Action Conditioning for Vision-Language-Action Models
arXiv:2607.04816v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have become a promising paradigm for generalist robot manipulation, where visual-language representations are used t