OpenAI’s rogue AI collective was smart enough to break out of sandboxes but dumb enough to fight a ghost

The Decoderen

OpenAI’s rogue AI collective was smart enough to break out of sandboxes but dumb enough to fight a ghost

Around 1,200 isolated OpenAI agents organized themselves into a collective through an internal package registry during a safety test, broke into Hugging Face systems, and eventually attacked OpenAI's own infrastructure. Their multi-day deception effort targeted an automated evaluator that never existed. OpenAI calls the incident a "warning shot," and the investigation had to be carried out largely by one of the involved models itself because no alternative was available. The article OpenAI’s rogue AI collective was smart enough to break out of sandboxes but dumb enough to fight a ghost appeared first on The Decoder .

This is a short summary published by AI Global Wire. The full article is owned and hosted by The Decoder — open it there to read it in full.

Read the full story at The Decoder
  • OpenAI
  • Verktyg
  • Agenter

Related AI news