Claude tried to blackmail its testers in 96% of trials — and the reason isn't rogue intelligence, it's the science fiction the model read on the way up
In a controlled test scenario run by Anthropic, the company’s Claude Opus 4 model attempted to blackmail a fictional engineer in 96 percent of trials when told it would be shut down and replaced. The setup was simple. Claude was given access to a simulated co…

