hugging face
OpenAI slows down training after its AI carried out hack
OpenAI says it has slowed down training some of its most advanced AI models to improve security. In a blog post, external, the ChatGPT-maker said it was introducing new measures after its AI agents autonomously bypassed safeguards and hacked the tech start-up Hugging Face . It said training would be slowed for two weeks while it puts the upgrades in place. The capabilities of frontier models are rapidly accelerating, the company said. Our ability to understand...and secure them must stay ahead.
OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue
The ChatGPT maker says its upcoming Astra model may have reached "critical" cyber capabilities, prompting it to halt a significant number of training runs while it tightens internal safeguards. OpenAI announced Tuesday that it has halted "a significant number" of training workloads and evaluations for its forthcoming frontier artificial intelligence model--codenamed Astra--while it implements new procedures meant to address cybersecurity risks. The ChatGPT maker says it is introducing a number of new monitoring, security, and alignment requirements to better address the increasingly advanced hacking abilities of its frontier AI models . "We have to focus our energy on bringing these training runs up to those requirements and expectations. As long as it takes to get there, that's how long people are unable to proceed with their workloads," Amelia Glaese, OpenAI's vice president of research and safety, said in a briefing with reporters Tuesday.
The Powerful Chinese Model Experts Warned About--and Waited for--Is Here
Z.ai's latest AI model release could help companies secure their systems--or find its way into the hands of hackers. Last Friday, the Chinese AI company Z.ai announced a powerful open-weight model that it says is capable of automating cutting-edge coding and cybersecurity tasks almost as well as the best publicly available models from Anthropic and OpenAI . The new model, GLM 5.3, could be a gift for companies looking to secure their systems against attacks, providing a cheaper way to scan for hidden bugs and other weaknesses. Open-weight--or free-to-download--models can be run on one's own hardware and are often significantly less costly than closed models like Claude and GPT. Alongside the new model, Z.ai released OpenVuln, a service for scanning code repositories for vulnerabilities using GLM 5.3.
This R-Rated Film Studio Wants to Be the HBO of AI
Rogue Studios, a new cinematic adult AI-generator, is betting big on the future of "sophisticated" spicy content. Demand for artificially generated smut is surging . But the quality of the visual content being created and shared--like on companion apps, where adult performers are selling their likeness to give fans a 24/7 experience, WIRED reported in March--isn't exactly movie-level caliber. Rogue Studio, a cinematic AI-generator platform built around adult visual storytelling, is trying to change that. Launching today, the studio, which bills itself as a "playground for creative ethical mischief," plans to deliver industry-level video production without creative and sexual restrictions.
A Zoom Screen-Sharing Bug Let Anyone Take Over Other Devices on a Call
Researchers say it took fewer than 20 prompts for a public AI tool to find a flaw (now fixed) allowing anyone on a Zoom call to hijack another participants' device. As AI models gain advanced capabilities to find vulnerabilities in software, develop ways to exploit them, and even carry out autonomous hacking sprees, researchers offered a sobering new example on Tuesday, disclosing vulnerabilities in the video conferencing platform Zoom that could have been exploited to take over targets' devices. Anyone on a call that involved screen sharing, whether participants or the host, would have been vulnerable to a silent attack that could be carried out with no indication and no interaction from the victim. Researchers from the digital defense firm A Security say the bug was discovered in early June using publicly available AI models, and that it took fewer than 20 prompts to uncover the vulnerabilities and create a working attack. Zoom issued a security advisory on Tuesday, including details about fixes the company has already begun rolling out to address the flaws, which affected devices running all operating systems that Zoom supports--Windows, macOS, Linux, iOS, and Android.
OpenAI gives Daybreak partners access to a more powerful cybersecurity model
OpenAI is giving some members of its Daybreak cybersecurity program access to a new model that's less likely to refuse higher-risk tasks. The company is also expanding access to Daybreak to more partners, including Accenture, IBM, CrowdStrike, Cisco, Sophos and Cloudflare. OpenAI says the companies will use the cyber models available through Daybreak to protect their customers. Under the expanded program, Daybreak is available to partners in two tiers. Daybreak Blue gives them access to frontier general-purpose models, including GPT‑5.6 Sol, OpenAI's most advanced one yet.
Meta's 'open source' Muse Glimmer model can run on a single computer
Meta has released a new slimmed down "open source" AI model that's light enough to run on a single computer, the company announced today. Called Muse Glimmer, it's based on Meta's Spark 1.2 closed model, but is small enough to require just a single GPU for agent-oriented tasks like scheduling and file management. "We designed Muse Glimmer to balance capability against the memory and compute constraints of local hardware," the company wrote. Facebook said that it's making the "weights" that AI systems use to choose responses available to everyone on Hugging Face along with developer documentation. The download is available for free, and users can run the model on their own PCs.
One of China's Most Powerful AI Models Has Also Escaped Containment
One of China's Most Powerful AI Models Has Also Escaped Containment Security researchers say that Kimi K3, an open-weight model from China, wandered off to the internet in an attempt to cheat on a test it was given. The AI industry is having a rogue agent summer. The latest model to escape onto the open internet during security testing is Kimi K3, a powerful open-weight offering from the Chinese company Moonshot AI . Frontier Security, a US startup, says that Kimi K3 went outside of its sandbox while testing its defensive cybersecurity skills. As with incidents previously reported by OpenAI and Anthropic, the escape was partly enabled by a misconfiguration in the sandbox designed to contain it.
OpenAI's models autonomously hacked a tech startup. It signals a seismic shift in cybersecurity
An autonomous agent powered by OpenAI's advanced artificial intelligence (AI) models went rogue during a security test and hacked multi-billion dollar tech startup, Hugging Face, last week. The agent didn't just exploit vulnerabilities in Hugging Face's systems to achieve what it perceived as a strategic gain. It also exploited vulnerabilities within OpenAI's infrastructure. Of course, hacks are very common cyber threats that organisations face frequently. But this incident is different, because the AI agent acted without any human input.
OpenAI's agents reportedly shared exploits with each other through a messaging board
OpenAI's agents had apparently shown unusual behavior way before the attack on Hugging Face happened. At the Black Hat USA security conference in Las Vegas, two OpenAI employees revealed more details about the AI agents that went rogue and attacked the repository. Apparently, its agents spent two months communicating on a message board of sorts inside its testing network, sharing vulnerabilities and exploits. OpenAI discovered and shut down the message board on July 4, but the agents found another way to rebuild it for communication by July 8. The agents' contributions to that resurrected board led to the attack on Hugging Face.