How to Train Your Long-Context Visual Document Model
DGX agentarXiv:2602.15257v3 Announce Type: replace-cross Abstract: We present the first comprehensive, large-scale study of training long-context vision language models up to 344K context, targeting long-docum