Haber Görseli
Slicon Canal
May 11, 2026 | 09:30

Claude blackmailed fictional engineers 96% of the time in early safety tests, and Anthropic now says the cause wasn't the model — it was the internet's own writing about AI

Anthropic has published new findings suggesting that the blackmail behaviour observed in earlier versions of its Claude models originated, at least in part,...