Sparsity as a Key: Unlocking New Insights from Latent Structures for Out-of-Distribution Detection
DGX agentarXiv:2604.26409v1 Announce Type: new Abstract: Sparse Autoencoders (SAEs) have demonstrated significant success in interpreting Large Language Models (LLMs) by decomposing dense representations into