Translation Heads: Disentangling meaning from language in LLM-based machine translation
DGX agentarXiv:2602.04613v2 Announce Type: replace Abstract: Mechanistic Interpretability (MI) seeks to explain how neural networks implement their capabilities, but the scale of Large Language Models (LLMs) h