Human-like Object Grouping in Self-supervised Vision Transformers
DGX agentarXiv:2603.13994v2 Announce Type: replace-cross Abstract: Vision foundation models trained with self-supervised objectives achieve strong performance across diverse tasks and exhibit emergent object s