Mechanistic Interpretability of Antibody Language Models Using SAEs
DGX agentarXiv:2512.05794v2 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) are a mechanistic interpretability technique that have been used to provide insight into learned concepts within la