SAME: A Semantically-Aligned Music Autoencoder
DGX agentarXiv:2605.18613v1 Announce Type: cross Abstract: Latent representations are at the heart of the majority of modern generative models. In the audio domain they are typically produced by a neural-audio