Same Concept, Different Directions: Cross-Modal Feature Heterogeneity in Sparse Autoencoders
DGX agentarXiv:2606.29888v1 Announce Type: cross Abstract: Vision-language models map images and text into a joint embedding space. However, these embeddings often entangle multiple semantic features, which li