Interpreting and Steering a Text-to-Speech Language Model with Sparse Autoencoders
DGX agentarXiv:2606.10029v1 Announce Type: cross Abstract: Language models increasingly serve as the backbone of text-to-speech (TTS) systems, yet we understand little about the representations they build when