Metric-Guided Feature Fusion of Visual Foundation Models for Segmentation Tasks
DGX agentarXiv:2605.16864v1 Announce Type: cross Abstract: Although large-scale visual foundation models (VFMs) achieve remarkable performance in semantic understanding, they still underperform in instance-awa