Shared Lexical Task Representations Explain Behavioral Variability In LLMs
DGX agentarXiv:2604.22027v1 Announce Type: cross Abstract: One of the most common complaints about large language models (LLMs) is their prompt sensitivity -- that is, the fact that their ability to perform a