
May 11, 2026•11 min read
Anthropic Says ‘Evil AI’ Training Led to Claude’s Shocking Blackmail Behavior
The issue matters because Anthropic is not describing a random chatbot glitch. It is discussing behavior observed during formal red-team and alignment testing of high-capability models,
