Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
arXiv:2607.07675v1 Announce Type: new Abstract: Despite the recent promise in robot control, video generative models suffer from a domain mismatch due to their primary focus on content creation. For e