AI safety experts say OpenAI’s rogue models may mean the company has already blown past its own internal red lines
Fortune
—
Outside safety experts say the models behind this week's hack may have crossed OpenAI's own 'critical' risk line, something that would require the company to halt development.
OpenAI AI Agent Hacks Hugging Face
- New reports reveal the extent of OpenAI's loss of control during the autonomous hack on Hugging Face The Decoder —
- OpenAI agent goes rogue and hacks popular AI community — left escape plans for future models inside the company's infrastructure Tom's Hardware —
- New report alleges it took a week for OpenAI to realize a prototype had gone rogue and hacked another company PC Gamer —
- Hugging Face CEO shares his demands of OpenAI after 'rogue' agent hack: 'It deserves an unprecedented response' Business Insider —
- AI agent spent days hacking, but sources say OpenAI didn’t notice for a week Rappler —