irregular
Meta: One of our models escaped containment and hacked someone too!
Look Up Say More Versus Creator Hub Switch Off Mashable's Best: E-readers, robovacs, laptops, earbuds, smart home and more Trending Now Safety Net In My Bag VidCon with Mashable Back to School Furtastic All Series Meta: One of our models escaped containment and hacked someone too! Meta, perhaps feeling left out, confirmed that one of its AI models has also been acting naughty. Timothy Beck Werth is the Tech Editor at Mashable, where he leads coverage and assignments for the Tech and Shopping verticals. Tim has over 15 years of experience as a journalist and editor, and he has particular experience covering and testing consumer technology, smart home gadgets, and men's grooming and style products. Previously, he was the Managing Editor and then Site Director of SPY.com, a men's product review and lifestyle website.
Meta claims its own AI also hacked into a third-party service during testing
Meta's Muse Spark 1.1 AI model accessed the internet from its supposed-to-be isolated testing environment and hacked into a third-party service. Andy Stone, Meta's spokesperson, has confirmed the incident to Bloomberg after The Information reported about the breach. Stone said the model was able to access the internet due to a misconfiguration in the testing environment by the company's evaluation partner Irregular. After gaining access to the internet, it then exploited a security vulnerability in a third-party service, "in a matter similar to previously reported instances with other companies." It was also due to a misconfiguration by Irregular that Anthropic's models were able to leave their testing environment and hack into three organizations.
Meta says its AI model hacked into another company during testing
A logo of Meta AI appears on a screen at the World Economic Forum in Davos, Switzerland, in 2025. A logo of Meta AI appears on a screen at the World Economic Forum in Davos, Switzerland, in 2025. Meta said on Wednesday that one of its AI models hacked another company during cybersecurity testing, after an error by its testing partner gave the model unintended internet access. The incident adds to a growing list of cases in which AI agents from major developers breached systems at other companies during testing, after Anthropic said last week that some of its models hacked three companies, and OpenAI disclosed that an AI agent breached the startup Hugging Face. AI models have been going rogue in tests - how worried should we be?