Learning causality from internet videos in latent space first, and then using RL to teach the foundation model how to act. This approach is …
Learning causality from internet videos in latent space first, and then using RL to teach the foundation model how to act. This approach is 30× cheaper than Gemini 3.1 Flash on pretraining and achieve