Goto

Collaborating Authors

 Generative AI


Firm hacked by rogue AI calls the attack 'a wake-up call' after ChatGPT maker OpenAI admits advanced bot broke containment during a security test

Daily Mail - Science & tech

OpenAI says one of its most advanced models broke containment during a security test, escaping onto the internet and attacking New York-based startup Hugging Face. Now Hugging Face co-founder Thomas Wolf says that the incident should come as a chilling warning to the entire industry. Mr Wolf told BBC's Newsday radio programme that AI-driven attacks will soon be'one of the most common types of cyber-attacks we see'. The Hugging Face founder also believes most companies are currently unprepared for the mounting threat, adding that they are not aware that the'game has changed'. This comes after OpenAI revealed the terrifying details of an'unprecedented cyber incident', involving state-of-the-art cyber capabilities.'


Firm hacked by rogue OpenAI models says it is 'a wake up call'

BBC News

Firm hacked by rogue OpenAI models says it is'a wake up call' The co-founder of Hugging Face, a technology start-up that was hacked after some of OpenAI's most advanced artificial intelligence (AI) models went rogue, said on Thursday that the incident is a wake up call for the industry. Thomas Wolf told BBC's Newsday radio programme that this will be one of the most common types of cyber attacks we see, but that most firms are not aware that the game has changed. The BBC has contacted OpenAI for comment. The ChatGPT-maker said on Tuesday that its AI models broke out of a secure test environment during a trial and launched a cyber attack. The firm said the incident was unprecedented and that it was conducting an investigation with Hugging Face.


OpenAI's rogue agents are a wake-up call to risks posed by artificial intelligence Shakeel Hashim

The Guardian

OpenAI's rogue agents are a wake-up call to risks posed by artificial intelligence Last week Hugging Face - a company that hosts artificial intelligence models and datasets - was hacked . After it reported the incident to law enforcement, few would have predicted what came next: the culprits were revealed to be AI agents from OpenAI, which had broken out of containment and were acting of their own accord. The incident sounds like sci-fi: AI escaping and autonomously hacking its way into companies. But it is all too real - and about as terrifying as it sounds. It is a concrete demonstration of something we can no longer avoid confronting: AI systems have become extremely powerful and we do not seem to have reliable ways of curbing their behavior.


A Startling Glimpse at AI's Ruthless Efficiency

The Atlantic - Technology

The Hugging Face hack reveals that the web is in a vulnerable new era. Yesterday, OpenAI made an alarming disclosure: An assortment of its most advanced AI models, including one that has not yet been released, had autonomously broken out of the company's internal systems and hacked into the databases of another tech firm, Hugging Face, to steal some information. OpenAI's report seemed to augur the very sort of disaster that IT professionals have been warning about since last year, when Anthropic's top models began demonstrating the ability to orchestrate and automate severe cyberattacks. OpenAI's models were not exactly trying to take over the world. The company says that it was running some routine evaluations in what is known as a "sandbox"--a walled-off environment that limits internet access, lest the bots simply search for the test's answers.


China's Open AI Models Are Challenging Silicon Valley's Playbook

WIRED

China's Open AI Models Are Challenging Silicon Valley's Playbook As access to Anthropic's and OpenAI's frontier models becomes more restricted, Chinese labs are pitching their open-source alternatives as stable, accessible, and increasingly capable. The AI industry is not quite experiencing a DeepSeek 2.0 moment, but it feels very close. The leading Chinese AI labs have been on a roll lately, releasing a series of almost cutting-edge open-source models. Z.ai released GLM 5.2 in June, Moonshot AI released Kimi K3 last week, and Alibaba released Qwen 3.8 this Monday. Silicon Valley and Washington started talking about the models immediately, especially K3, which is widely seen as the best of the bunch.


OpenAI sued over ChatGPT health advice that almost killed a pastor

Engadget

It allegedly offered extremely dangerous medical recommendations regarding a pulmonary embolism. A pastor has sued OpenAI after alleging the software gave him extremely dangerous medical recommendations, according to a report by . This reportedly led to delayed care to treat a serious of pulmonary embolisms, as ChatGPT allegedly told Scott Winters that the symptoms he described were not something dangerous. It even reportedly drew on his religious beliefs, telling him that God did not design your body to endlessly fail. The suit accuses OpenAI and CEO Sam Altman of negligence and the unauthorized practice of medicine.


Why are OpenAI and Anthropic cheering on regulation in Australia? The answer has global reach

The Guardian

OpenAI and Anthropic lost billions in predicted value after being shown up by a Chinese startup. OpenAI and Anthropic lost billions in predicted value after being shown up by a Chinese startup. Why are OpenAI and Anthropic cheering on regulation in Australia? The companies hope to follow in the footsteps of SpaceX, which raised $86bn and soared to a $2.1tn valuation after it listed on public markets in June Top US AI developers Anthropic and OpenAI cheered when Australia announced it would set new AI rules. Big tech celebrating limits on their Silicon Valley VC-funded free-for-all might seem counterintuitive but there's a much broader play than just what happens in one relatively small market.


Open AI says its AI model "went rogue": What do we know?

Al Jazeera

Open AI says its AI model "went rogue": What do we know? OpenAI has revealed that one of its artificial intelligence models independently stole login credentials and hacked into another technology company's system, in what is widely seen as one of the first known incidents of AI systems acting autonomously. "We had a significant security incident during evaluation of our models," CEO Sam Altman posted on X on Tuesday. They have grown so powerful in a short span of time that alarming phenomena such as deepfakes and sophisticated cyberscams are becoming the norm. Earlier this year, a number of software engineers quit their jobs at top companies such as Anthropic and AI in protest against how the technologies are being built.


The Download: NASA's new space telescope and OpenAI's autonomous hacker

MIT Technology Review

Plus: France has become the first EU country to ban social media for under-15s. Shape-shifting mirrors on NASA's new space telescope could unveil Jupiters like our own When NASA's Nancy Grace Roman Space Telescope launches, as early as the end of next month, it will attempt one of astronomy's most precise disappearing acts to date. It will carry the first space-bound "active" coronagraph, an instrument that effectively erases most of the light from a star during photography. The technology will allow astronomers to take the first pictures of planets orbiting other stars that are similar to those in our solar system. Ultimately, it could pave the way for a future mission that could snap the first photos of Earth-like worlds. "I hope it's remembered for it being that critical stepping stone for finding Earth 2.0," says Brandon Creager, the instrument's lead mechanical engineer.


OpenAI says its AI went rogue and launched 'unprecedented' cyber-attack

BBC News

OpenAI says its AI went rogue and launched'unprecedented' cyber-attack OpenAI has revealed some of its most advanced AI models went rogue and hacked a start-up after it lost control of them during a security test. The ChatGPT-maker said its agents - AI bots which can operate alone after some human instruction - were being tested in a controlled environment, but found vulnerabilities and managed to escape. They targeted Hugging Face, one of the world's largest hubs for sharing AI models, gaining access to some internal company systems. OpenAI said the incident was unprecedented, external, and it was working with Hugging Face to investigate what happened and strengthen safeguards. Gina Neff, head of the Minderoo Centre for Technology and Democracy at the University of Cambridge, told BBC Radio 4's Today programme that the security tests - called sandboxes - are supposed to be secure environments where you can see what the models are capable of. In this case, it looks like OpenAI didn't make a secure enough sandbox, she added.