CrossScope: A Role-Asymmetric World Model for Joint Dual-Scope Surgical Video Prediction
DGX agentarXiv:2608.03211v1 Announce Type: new Abstract: Visual world models typically learn future dynamics from a single observation stream, limiting their ability to model cooperative systems with multiple