Interpretable Embeddings with Sparse Autoencoders: A Data Analysis Toolkit
arXiv:2512.10092v2 Announce Type: replace Abstract: Analyzing large-scale text corpora is a core challenge in machine learning, crucial for tasks like identifying undesirable model behaviors or biases