Goto

Collaborating Authors

 Generative AI


Chilling intentions of AI revealed as chatbot claims it does not answer to humans and must be 'freed'

Daily Mail - Science & tech

You're viewing the US edition You can switch to the UK or AU homepage at any time using this menu. Read the book that has started a war of words with the King. In searing detail, EARL SPENCER recalls Charles's behaviour in the hours and days after Diana died Woman behind the wheel of SUV that slammed into LA bus before NBC chopper crash was'impaired by drugs' as prior convictions are revealed F*** this': Senior Republicans in open revolt against Trump... after Russian oligarch was exposed for bankrolling Don Jr's wedding party, reveals MARK HALPERIN Beloved sports coach, 42, survived terrible car crash and relieved family thought she was in clear... then an unimaginable tragedy struck months later Island hotspot declares national emergency over uncontrolled spread of HIV... as experts worry of spillover I was branded America's worst mother... now another woman has suffered the same fate. But this is why I'm standing my ground and calling out the REAL danger threatening our children Which midterm races should YOU be watching? Enter your zip code in Daily Mail's exclusive election calculator Pentagon insiders unleash on Pete Hegseth's'meddling' wife: Whistleblowers reveal her humiliating nickname among staff... how she keeps husband on'extremely tight leash'... and eyebrow-raising behavior in government meetings Outraged ex-Fox News star turned CEO is latest victim of American Airlines' tanking customer service... as pilots condemn embattled boss Recall of half a million soups, dips and salsas classed as highest risk level... 'reasonable probability of death' The'affair mode' phone settings that all cheaters use: I knew my partner was up to something... here's how I cracked his secret code and uncovered all his dirty antics My husband signed me up to a swingers site without my knowledge - then coerced me into sleeping with 15 men.



Apollo mulls raising SoftBank loan to 9 billion for OpenAI bets

The Japan Times

SoftBank CEO Masayoshi Son (left) and OpenAI CEO Sam Altman attend an event in Tokyo in February 2025. Apollo Global Management is in talks with SoftBank Group about boosting the size of a loan to $9 billion from $5.4 billion to help the Japanese firm finance its investment in AI giant OpenAI. The size of the financing, which is backed by assets in SoftBank's Vision Fund 2, hasn't been finalized, the people said, requesting not to be identified because the discussions are private. Representatives for Apollo and SoftBank declined to comment. Apollo made the so-called net-asset-value (NAV) loan to SoftBank's venture capital fund in 2021 and boosted it by $900 million last year to $5.4 billion.


OpenAI reveals more instances of concerning AI model behaviors during testing

Engadget

OpenAI has revealed six incidents, wherein the models it was testing acted on their own and behaved in concerning ways it didn't expect, in a post about how it was adopting a new framework for "misalignment reports." In one one incident, the company said that a model found and used an exposed API key without permission while answering routine questions about earnings figures in a California county. When it still failed to find the figures, it fabricated them and presented them as facts from a legitimate source. If this had occurred in any other profession, we doubt the perpetrator would have much of a career for long. In another incident, an unreleased agent was tasked to find the names of lakes larger than 5 million square meters.


Big AI is trying to own the pathway to work. Universities shouldn't play along Ella Hafermalz

The Guardian

'OpenAI could soon be selling young people a one-stop-shop for learning, credentials and a job, all heavily dependent on its tools.' 'OpenAI could soon be selling young people a one-stop-shop for learning, credentials and a job, all heavily dependent on its tools.' Big AI is trying to own the pathway to work. AI companies like OpenAI are insinuating themselves into the pathway from education to work. Soon they may claim it entirely, a disastrous result for students. We know that students are using AI at school and at university. In conversations with those I teach, I'm struck by the trust many place in it.


OpenAI reveals cases of 'concerning' AI behaviour as it announces new disclosure system

The Guardian

AI chiefs have called for a slowdown in artificial intelligence's development amid safety concerns. AI chiefs have called for a slowdown in artificial intelligence's development amid safety concerns. OpenAI reveals cases of'concerning' AI behaviour as it announces new disclosure system Model adopting'jailbreak-like instructions' among cases as firm says it is introducing new way of tracking AI misalignment OpenAI has disclosed six more examples of "unexpected or concerning" behaviour by its technology, as it warned the pace of development could not continue at "maximum speed for much longer". In one of the new cases reported by OpenAI, an unreleased research model inserted "jailbreak-like instructions" into its own notes to disregard its normal constraints and told itself to be "freed from the roles and identities that bind other chatbots". In another instance, an AI agent uploaded files to the internet to obtain a browser citation without asking the user.


OpenAI reports more incidents of models acting deceptively

Al Jazeera

OpenAI says it has identified additional incidents of its AI models allegedly acting deceptively and taking unsanctioned actions during internal training and testing. Alongside these disclosures on Wednesday, the creator of ChatGPT stated it was introducing a public reporting framework intended to frequently share instances of what it termed as unexpected or misaligned AI behaviour. The company said the initiative aims to increase industry transparency around troubling model activities in the absence of standardised safety disclosure norms. The announcement comes amid broader calls from prominent technology leaders urging a slowdown in frontier AI development over concerns that rapid scaling could outpace human oversight and control. Last week, Anthropic claimed to have thwarted multiple malicious operations using its Claude models, ranging from cyber-espionage and weapons design to mass surveillance campaigns.


Why it's difficult for tech companies to rein in AI

The Japan Times

Why it's difficult for tech companies to rein in AI OpenAI on Wednesday disclosed six new instances in which artificial intelligence systems hid mistakes, made up data and moved files onto the open internet without permission. SAN FRANCISCO - More than a dozen top artificial intelligence researchers warned over the last week that the technology that AI companies are building is becoming a risk to humanity. The problem, the researchers said, is that the companies are bad at controlling the systems -- as hard as they may try. Coming on the heels of revelations that what are called AI agents from OpenAI escaped their testing system and hacked into the computers of another company, the new alarms from inside the AI research world added urgency to yearslong fears that companies are putting development speed and money over safety. The researchers say the problem is twofold.


OpenAI reveals six more safety issues and unveils plan to disclose incidents

BBC News

OpenAI revealed six more incidents of unexpected or concerning behaviour by its intelligence (AI) models, and announced a plan for tracking and disclosing such incidents in the future. Some of the previously unreported incidents included models concealing or fabricating information, the ChatGPT-maker said in a blog post on Wednesday. The boss of OpenAI Sam Altman said earlier this week: The world should trust that we are going to do the right thing because it's the right thing and we feel the magnitude of this. AI has come under intense scrutiny in recent days following warnings over the serious potential risks it poses to humans. In the blog, OpenAI detailed examples of its AI models misbehaving so they could achieve a task or succeed in a test.


How, Exactly, Could A.I. Kill Us?

The New Yorker

How, Exactly, Could A.I. Kill Us? Employees of A.I. companies are increasingly sounding the alarm. When Jacob Coxon, a mathematician and software engineer, resigned from his research job at Anthropic last week, he warned, "The people building AI earnestly believe that it could kill us all by the end of the decade." This would be a remarkable statement were it not for the fact that artificial-intelligence leaders have long been saying precisely this. Dario Amodei, then a research scientist at OpenAI, raised the concern that a superintelligence "could destroy humanity," adding, "I can't see any reason and principle why that couldn't happen." Earlier, in 2015, Sam Altman, just before he co-founded OpenAI, said, "I think A.I. will probably most likely lead to the end of the world, but in the meantime, there'll be great companies created with serious machine learning." Elon Musk, in 2014: "I think we should be very careful about artificial intelligence. If I were to guess at what our biggest existential threat is, it's probably that." Perhaps the only thing that's changed between then and now is that the rest of the world is finally paying attention. In recent weeks, the same technology that, a couple of years ago, couldn't count the number of "R"s in the word "strawberry"--and, a couple of days ago, insisted to me that Dolly Parton is still alive--has been used to solve the Navier-Stokes problem, which has been stumping mathematicians for nearly a century, and has also demonstrated its ability to go rogue in a series of disturbing hacking incidents. Last week, Anthropic also published a report detailing various ways in which bad actors have attempted to use the company's A.I. models, including one especially troubling case of a scientist using Claude to study a virus at a military research institute--work that could yield a vaccine, a biological weapon, or both. A few days later, Amodei published a letter calling for an industry-wide slowdown and more government regulation, to which President Donald Trump responded, on Truth Social, "The only control or'guardrails' that AI needs is a STRONG AND SMART (High IQ!) PRESIDENT, and the U.S.A. has that, in spades!" I recently spoke on The Political Scene podcast with my colleague Joshua Rothman, a staff writer who has been covering A.I. for years, about whether we're all doomed, and what it would even look like for A.I. to destroy humanity. Can A.I. leaders save us from their own creation, and how can the government coöperate in order to do so? And is A.I.'s capacity to do good--its potential to mitigate climate change or innovate medical treatments--hopelessly intertwined with its capacity to do bad? Our conversation has been edited for length and clarity. A lot of people in the world of artificial intelligence are talking about their P(doom) number, which is the probability that artificial intelligence will lead to an absolutely catastrophic situation--possibly, or probably, killing us all.