ExToken: Structured Exploration for Efficient Vision-Language-Action Reinforcement Fine-tuning
arXiv:2607.12931v1 Announce Type: new Abstract: Reinforcement Learning (RL) has demonstrated significant potential for improving Vision-Language-Action (VLA) models on complex manipulation tasks. Howe