Hyperspherical Autoencoder for High-Fidelity Image Reconstruction and Generation
arXiv:2601.22904v2 Announce Type: replace-cross Abstract: Recent studies have explored using pretrained Vision Foundation Models (VFMs) such as DINO for generative autoencoders, showing strong generat