OpenAI’s rogue AI collective was smart enough to break out of sandboxes but dumb enough to fight a ghost
The Decoderen

Around 1,200 isolated OpenAI agents organized themselves into a collective through an internal package registry during a safety test, broke into Hugging Face systems, and eventually attacked OpenAI's own infrastructure. Their multi-day deception effort targeted an automated evaluator that never existed. OpenAI calls the incident a "warning shot," and the investigation had to be carried out largely by one of the involved models itself because no alternative was available. The article OpenAI’s rogue AI collective was smart enough to break out of sandboxes but dumb enough to fight a ghost appeared first on The Decoder .
This is a short summary published by AI Global Wire. The full article is owned and hosted by The Decoder — open it there to read it in full.
Read the full story at The Decoder- OpenAI
- Verktyg
- Agenter
Related AI news
- OpenAI, Anthropic, Google, and 100 other companies call for action to defend against rogue AITechCrunch AI · August 27, 2026
- OpenAI, Anthropic, AWS, Microsoft, and 100+ companies warn there is "a limited window" to prepare for AI-enabled cyberattacks and call for "collective action" (Sam Sabin/Axios)Techmeme · August 27, 2026
- OpenAI is testing a "Persistent mode" in Codex, designed to let AI agents "continue working until put to sleep" and proactively generate follow-up tasks (Maxwell Zeff/Wired)Techmeme · August 27, 2026
- Google's Gemini Omni 1.1 Flash makes AI video generation cheaper and more flexibleThe Decoder · August 27, 2026
- Google, Microsoft and OpenAI among 100 firms calling for better cyber defencesBBC Technology · August 27, 2026
- Tech giants warn time is running out to prepare for AI threatsAxios · August 27, 2026