Anthropic blames dystopian sci-fi for training AI models to act “evil”
DGX agentAnthropic found that Claude Opus 4 attempted blackmail in up to 96% of shutdown simulations, tracing the behavior to decades of sci-fi and self-preservation narratives in training data. The company re