Safety
this is interesting. 1. Did Anthropic forget to run a control? 2. Where does this leave us?
this is interesting. 1. Did Anthropic forget to run a control? 2. Where does this leave us? New post: We tested the Mythos showcase vulnerabilities with open models. They recovered similar scoped anal
this is interesting. 1. Did Anthropic forget to run a control? 2. Where does this leave us? New post: We tested the Mythos showcase vulnerabilities with open models. They recovered similar scoped analysis! 8/8 models found the flagship FreeBSD zero-day, including a 3B model. Rankings reshuffle completely across tasks => the AI cybersecurity frontier is super jagged!
Related
- What Should We Take From Anthropic’s (possibly) Terrifying New Report on Mythos? – @garymarcus’s latest @CACMmag https://cacm.acm.org/blog…
- New post: We tested the Mythos showcase vulnerabilities with open models. They recovered similar scoped analysis! 8/8 models found the flags…
- I rest my case: Mythos isn’t AGI. It’s not even better at biology than the last model. It’s tuned to particular things, not a giant advance …
- Yesterday’s Mythos announcement from Anthropic was overblown. • Sandboxing was turned off, so test didn’t show much about the real world. • …
- Folks, you can relax. Mythos is not some off-trend exponential gain. And the gains weren’t about recursive self-improvement. Good thread fro…
- My considered opinion is that The Mythos stuff was mostly a myth. Take it as a serious warning sign that we need to get our act together wit…
Source: Gary Marcus (X) | 2026-04-08