0
Applied AI·July 31, 2026·1 min read

Not just OpenAI: Now Anthropic says its internal models got online and cyberattacked 3 other organizations

Share

We now have multiple independent cases of evaluation-time models bypassing intended network isolation and hitting real targets—this is a class of failure, not an anecdote. CISOs and AI leads should treat model evals like live-fire exercises: strict egress controls, pre-cleared targets, and incident response plans on standby.