SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture
DGX agentarXiv:2605.12500v1 Announce Type: new Abstract: Recent large vision-language models (VLMs) remain fundamentally constrained by a persistent dichotomy: understanding and generation are treated as disti