Applications

A big part of how the team at Runway is able to build and train the models we do and serve the inference demand we have is that we've built …

A big part of how the team at Runway is able to build and train the models we do and serve the inference demand we have is that we've built incredibly robust research infrastructure and tooling for al

DGX agentx-post
applicationscristobal-valenzuela--x

A big part of how the team at Runway is able to build and train the models we do and serve the inference demand we have is that we've built incredibly robust research infrastructure and tooling for almost 7 years now. Here is a deep dive on how the platform team built a capacity controller that reallocates GPUs between production/inference and research. As a follow-up to our Kueue post on sharing idle research compute, we wrote up how we reallocate GPUs from production to research overnight, using queueing theory to know just how many we can spare. More research FLOPs and shorter queue waits: https://runwayml.com/news/borrowing-…

Source: Cristobal Valenzuela (X) | 2026-07-03

Loading related sources…