Can Cross-Layer Transcoders Replace Vision Transformer Activations? An Interpretable Perspective on Vision
arXiv:2604.13304v1 Announce Type: new Abstract: Understanding the internal activations of Vision Transformers (ViTs) is critical for building interpretable and trustworthy models. While Sparse Autoenc