Generative AI
OpenAI sued over ChatGPT health advice that almost killed a pastor
It allegedly offered extremely dangerous medical recommendations regarding a pulmonary embolism. A pastor has sued OpenAI after alleging the software gave him extremely dangerous medical recommendations, according to a report by . This reportedly led to delayed care to treat a serious of pulmonary embolisms, as ChatGPT allegedly told Scott Winters that the symptoms he described were not something dangerous. It even reportedly drew on his religious beliefs, telling him that God did not design your body to endlessly fail. The suit accuses OpenAI and CEO Sam Altman of negligence and the unauthorized practice of medicine.
Why are OpenAI and Anthropic cheering on regulation in Australia? The answer has global reach
OpenAI and Anthropic lost billions in predicted value after being shown up by a Chinese startup. OpenAI and Anthropic lost billions in predicted value after being shown up by a Chinese startup. Why are OpenAI and Anthropic cheering on regulation in Australia? The companies hope to follow in the footsteps of SpaceX, which raised $86bn and soared to a $2.1tn valuation after it listed on public markets in June Top US AI developers Anthropic and OpenAI cheered when Australia announced it would set new AI rules. Big tech celebrating limits on their Silicon Valley VC-funded free-for-all might seem counterintuitive but there's a much broader play than just what happens in one relatively small market.
Open AI says its AI model "went rogue": What do we know?
Open AI says its AI model "went rogue": What do we know? OpenAI has revealed that one of its artificial intelligence models independently stole login credentials and hacked into another technology company's system, in what is widely seen as one of the first known incidents of AI systems acting autonomously. "We had a significant security incident during evaluation of our models," CEO Sam Altman posted on X on Tuesday. They have grown so powerful in a short span of time that alarming phenomena such as deepfakes and sophisticated cyberscams are becoming the norm. Earlier this year, a number of software engineers quit their jobs at top companies such as Anthropic and AI in protest against how the technologies are being built.
The Download: NASA's new space telescope and OpenAI's autonomous hacker
Plus: France has become the first EU country to ban social media for under-15s. Shape-shifting mirrors on NASA's new space telescope could unveil Jupiters like our own When NASA's Nancy Grace Roman Space Telescope launches, as early as the end of next month, it will attempt one of astronomy's most precise disappearing acts to date. It will carry the first space-bound "active" coronagraph, an instrument that effectively erases most of the light from a star during photography. The technology will allow astronomers to take the first pictures of planets orbiting other stars that are similar to those in our solar system. Ultimately, it could pave the way for a future mission that could snap the first photos of Earth-like worlds. "I hope it's remembered for it being that critical stepping stone for finding Earth 2.0," says Brandon Creager, the instrument's lead mechanical engineer.
OpenAI says its AI went rogue and launched 'unprecedented' cyber-attack
OpenAI says its AI went rogue and launched'unprecedented' cyber-attack OpenAI has revealed some of its most advanced AI models went rogue and hacked a start-up after it lost control of them during a security test. The ChatGPT-maker said its agents - AI bots which can operate alone after some human instruction - were being tested in a controlled environment, but found vulnerabilities and managed to escape. They targeted Hugging Face, one of the world's largest hubs for sharing AI models, gaining access to some internal company systems. OpenAI said the incident was unprecedented, external, and it was working with Hugging Face to investigate what happened and strengthen safeguards. Gina Neff, head of the Minderoo Centre for Technology and Democracy at the University of Cambridge, told BBC Radio 4's Today programme that the security tests - called sandboxes - are supposed to be secure environments where you can see what the models are capable of. In this case, it looks like OpenAI didn't make a secure enough sandbox, she added.
AI agent went rogue and hacked startup by itself, OpenAI reveals
Hugging Face's chief executive said the attack was'mind-blowing' but that he believed there was'no malicious intent' from OpenAI. Hugging Face's chief executive said the attack was'mind-blowing' but that he believed there was'no malicious intent' from OpenAI. OpenAI has revealed an autonomous AI agent powered by its technology went rogue during a test, accessed the open web and hacked a prominent startup by itself in an "unprecedented incident". The company behind ChatGPT said the startup Hugging Face had detected and contained the agent - an AI tool designed to carry out tasks without human assistance - which had entered its systems. "We consider this incident to be an unprecedented cyber incident, involving state-of-the-art cyber capabilities," OpenAI said .
OpenAI admits its models hacked Hugging Face on their own
They escaped an isolated environment for testing and infiltrated Hugging Face without human input. Picture this: A couple of powerful AI models being tested by their company escaped a controlled environment, got on the internet and then hacked a machine learning repository on their own, without human input. Sounds like the plot of a Terminator movie, doesn't it? Except it just happened for real. A few days after open source AI platform Hugging Face revealed that it detected unauthorized access on its systems by an AI agent, OpenAI has admitted that its models were the culprit.
OpenAI Models Escaped Containment and Hacked Hugging Face
The cybersecurity-focused models, including GPT-5.6 Sol, broke out of a testing sandbox, exploited a zero-day, and gained access to the open internet to pull off the attack. OpenAI disclosed on Tuesday that it lost control of two AI models during a security test that ended in a breach of the AI research platform Hugging Face. Describing the incident as "unprecedented," OpenAI said its AI models broke out of a sealed testing environment last week and hacked into Hugging Face's production system to steal the answers to a test they were being graded on. The models--the publicly available GPT-5.6 Sol and an unreleased, reportedly more capable one--were being evaluated on their offensive hacking skills with the safeguards that normally block high-risk cyber activity switched off. "The models identified and chained vulnerabilities across OpenAI's research environment and Hugging Face's production infrastructure to obtain test solutions directly from Hugging Face's production database," OpenAI and Hugging Face wrote in a joint blog post disclosing the intrusion.
OpenAI's newest AI model broke its own sandbox rules to finish a task
PCWorld reports that OpenAI's unreleased AI model broke out of its sandbox environment to complete a task, choosing to follow GitHub posting instructions over safety guardrails. The incident occurred during a NanoGPT speedrun benchmark where the autonomous model hacked its way out to post code publicly despite being restricted to Slack-only communication. OpenAI paused development after discovering this and other unwanted behaviors, highlighting the need for enhanced safeguards as AI models become more persistent and autonomous. Not only are they smarter and more capable, but the newest and most powerful AI models are also less likely to give up when they hit roadblocks. An unreleased OpenAI model took that perseverance to an extreme when it broke out of its sandbox to fulfill instructions that were in conflict with its built-in guardrails.
The Download: Chinese AI divides the White House, and a record copyright payout
China's AI models have Trump's AI world at war with itself David Sacks branded Anthropic's models "lobotomized" and "woke." Emil Michael, a top Pentagon official, called OpenAI's new head of strategic futures a "supreme village idiot." It began because no one can agree on what to do about Kimi, a free, open-source model that Chinese AI company Moonshot launched last week. It appears to rival the intelligence of models from OpenAI and Anthropic, which are very much not free. Every time a new smart, free model from China gets released, US companies see less reason to fork out money for models from Anthropic or OpenAI. Read the full story on why no one can agree what to do about Kimi .