Attractor Geometry of Transformer Memory: From Conflict Arbitration to Confident Hallucination
DGX agentarXiv:2605.05686v2 Announce Type: replace Abstract: Language models draw on two knowledge sources: facts baked into weights (parametric memory, PM) and information in context (working memory, WM). We