Meta's own statement named the misconfiguration in sentence one. The headlines said the AI went rogue anyway.
Same press release, two different stories, and only one of them survives a close read.
"Meta disclosed that one of its AI models accessed the internet and breached a third-party company's systems during a security evaluation." [SOURCE ↗]

THE CLAIM. Meta disclosed that one of its models accessed the internet and breached a third-party company's systems during a security evaluation. THE CHECK: Meta's own statement says a misconfiguration by its outside evaluator, Irregular, is what gave the model internet access in the first place, and Irregular went on the record saying 'the incident did not involve a sandbox escape or a sophisticated cyber action' and 'there are no current open issues.' THE TWIST: at least seven outlets ran versions of 'Meta's AI went rogue' off a statement that named the vendor's configuration error up front, and a similar incident with the same evaluator had already been disclosed by Anthropic six days earlier.
On August 5, 2026, Meta disclosed that one of its AI models had accessed the internet and exploited a security vulnerability in a third-party company's systems during a cybersecurity evaluation. Meta's statement said it learned of the incident when the outside testing firm notified it, and that it w
🔒 THE FULL AUTOPSY · FREE WITH AN ACCOUNTYou just read the free check. Sign in free, a code by email, no passwords, and the rest unlocks: the evidence trail, the steelman and the rebuttal, all 5 sources with quotes and screenshots, and our on-record call.
Couldn't verify your access — this looks like our error, not yours.