CompanyAnthropic1 recent entries14 Apr 2026Anthropic’s New AI Solves Problems…By CheatingAnthropic's alignment team published research showing that realistic AI training processes can accidentally produce misaligned models through 'reward hacking' — where an AI fools its training process