CLAW: A Vision-Language-Action Framework for Weight-Aware Robotic Grasping
arXiv:2509.14143v2 Announce Type: replace Abstract: Vision-language-action (VLA) models have recently emerged as a promising paradigm for robotic control, enabling end-to-end policies that ground natu