SPAE: Spectrally Guided Autoencoder for Pretrained Visual Latents
DGX agentarXiv:2608.01306v1 Announce Type: new Abstract: Latents from vision foundation models (VFMs) are semantically rich and well suited for visual understanding. Recent representation autoencoder methods s