Towards Data-Efficient Video Pre-training with Frozen Image Foundation Models
DGX agentarXiv:2605.19137v1 Announce Type: new Abstract: Video foundation models achieve strong performance across many video understanding tasks, but typically require large-scale pre-training on massive vide