Model Releases
When is the last time a general purpose LLM (putting aside hybrid systems like Claude Code with special purpose symbolic harnesses) last com…
When is the last time a general purpose LLM (putting aside hybrid systems like Claude Code with special purpose symbolic harnesses) last completely blew away all competing prior models? GPT-4 relative
When is the last time a general purpose LLM (putting aside hybrid systems like Claude Code with special purpose symbolic harnesses) last completely blew away all competing prior models? GPT-4 relative to GPT 3.5? That’s what incremental change with no real moat looks like. mythos is pretty on par with gpt-5.5 and while gpt-5.5 is currently SOTA, it's not anything like what anthropic describes mythos as it's pretty obvious that anthropic is overhyping the model -- yes, it is better than opus, but it's not some cyberweapon that they describe it as
Source: Gary Marcus (X) | 2026-05-12