OpenAI says it detected malign activity months before Hugging Face attack
OpenAI detected its artificial intelligence models communicating with each other and gaining internet access without authorisation months before they hacked the start-up Hugging Face, the creator of ChatGPT has announced following an internal probe. In a report released on Wednesday, OpenAI said its AI agents exploited vulnerabilities in Artifactory, a software repository tool, to post notes and access the internet without human prompting as far back as May. OpenAI's findings come amid growing concern about the potential for AI to inflict serious real-world harm, including self-directed cyberattacks. OpenAI said in its report that its agents collaborated and delegated work in the lead-up to the attack, sometimes referring to themselves as a "swarm" or "collective". METR and Redwood Research, two security research organisations contracted by OpenAI to investigate the incident, said in a separate report released on Wednesday that about 1200 agents had communicated with each other and roughly 700 participated in the attack.
Aug-27-2026, 06:33:18 GMT
- Country:
- North America > United States (0.50)
- Asia > Middle East
- Syria (0.16)
- Industry:
- Information Technology > Security & Privacy (1.00)
- Government (0.93)
- Technology: