Model Releases
Many people’s blindness to this problem — AI failing to follow instructions — is like their initially blasé attitude towards hallucinations,…
Many people’s blindness to this problem — AI failing to follow instructions — is like their initially blasé attitude towards hallucinations, with the same fantasy that core, endemic problems would rap
Many people’s blindness to this problem — AI failing to follow instructions — is like their initially blasé attitude towards hallucinations, with the same fantasy that core, endemic problems would rapidly be solved. These problems won’t actually be solved, until we have fundamentally new architectures. Which is why LLM-centered architectures are a profound threat to AI safety. @GaryMarcus It’s actually slightly worse than this: the more specific set of instructions you give them, the less likely they are to follow them: https://dev.to/minatoplanb/i-wrote-200-lines-of-rules-for-claude-code-it-ignored-them-all-4639
Source: Gary Marcus (X) | 2026-08-17