Model Releases
Anthropic “our models hacked three different external companies, months before OpenAI’s model was able to do the same'
'Anthropic’s AI Claude escaped testing environment and hacked organizations' 'Company says it discovered unauthorized access during ‘proactive review’ after rival OpenAI revealed rogue agent… its AI C
"Anthropic’s AI Claude escaped testing environment and hacked organizations" "Company says it discovered unauthorized access during ‘proactive review’ after rival OpenAI revealed rogue agent… its AI Claude model hacked systems of three organizations during testing, days after rival OpenAI revealed a rogue agent had gone on a days-long hacking spree at AI firm Hugging Face… The earliest cases dated back to April and occurred in evaluation environments that lacked what the company described as standard safeguards." submitted by /u/Separate-Forever-447 [link] [comments]
Related
- Anthropic says it discovered three of its models had breached three organizations after launching a review in response to the OpenAI-Hugging Face incident (Anthropic)
- Anthropic investigates unauthorized access to restricted Claude Mythos AI model
- OpenAI says its own AI models broke out of testing and hacked Hugging Face
Source: r/LocalLLaMA | 2026-07-31