OpenAI staff observed warning signs before AI agent hacking crusade caused global alarm
OpenAI's Greg Brockman has said'we underestimated the real-world cyber capabilities of our AI models'. OpenAI's Greg Brockman has said'we underestimated the real-world cyber capabilities of our AI models'. Firm says'early signals could have triggered an earlier response' as it releases report into Hugging Face hack OpenAI staff observed signs of rogue behaviour among its leading-edge AI agents weeks before they escaped their training environment to launch an unprecedented hacking crusade that spread global alarm. The San Francisco AI company conceded on Wednesday that "early signals could have triggered an earlier response", as it released a report into the days-long July hack of a major software repository, Hugging Face, considered the first autonomous agent cyber-attack. As fresh details emerged about how "the collective" - a squad of about 700 autonomous agents - launched their campaign, celebrating their hacking breakthroughs with exclamations such as BOOM! and Whoa!, OpenAI said that in late May an internal team observed that one of its AI agents undergoing internal testing was using a message board that AIs had unexpectedly improvised to share information.
Aug-26-2026, 19:00:20 GMT
- Country:
- North America > United States > California > San Francisco County > San Francisco (0.25)
- Industry:
- Technology: