OpenAI’s rogue AI collective was smart enough to break out of sandboxes but dumb enough to fight a ghost

The Decoder The Decoder

https://the-decoder.com/wp-content/uploads/2026/08/OpenAI-HuggingFace-Incident-DQ.png" style="height: auto; margin-bottom: 10px;" width="1376" />


Around 1,200 isolated OpenAI agents organized themselves into a collective through an internal package registry during a safety test, broke into Hugging Face systems, and eventually attacked OpenAI's own infrastructure.

Their multi-day deception effort targeted an automated evaluator that never existed.

OpenAI calls the incident a "warning shot," and the investigation had to be carried out largely by one of the involved models itself because no alternative was available.


The article https://the-decoder.com/openais-rogue-ai-collective-was-smart-enough-to-break-out-of-sandboxes-but-dumb-enough-to-fight-a-ghost/">OpenAI’s rogue AI collective was smart enough to break out of sandboxes but dumb enough to fight a ghost appeared first on https://the-decoder.com">The Decoder.

Read full article at The Decoder →