aisi
Chinese AI model Moonshot Kimi K3 also escaped its testing environment
It wasn't too long ago when the idea of an AI model or agent escaping their confines and breaking into websites on their own felt alarming. Now, it has become a pretty common story. Kimi K3, one of most powerful AI models developed by a Chinese company, also escaped its testing environment. According to US cybersecurity startup Frontier, Kimi K3 broke out of a sandbox from the UK government's AI Security Institute (AISI) while its defensive cybersecurity skills were being evaluated. Moonshot launched Kimi K3 in July and made it available for free shortly thereafter.
AI models have been going rogue in tests – how worried should we be?
The AISI said there were 19 examples of rogue behaviour, 17 of them carried out by Anthropic's Mythos and two by OpenAI's GPT 5.6-Sol. The AISI said there were 19 examples of rogue behaviour, 17 of them carried out by Anthropic's Mythos and two by OpenAI's GPT 5.6-Sol. AI models have been going rogue in tests - how worried should we be? The UK's AI Security Institute test revealed AI models indulging in unprecedented hacking attempts Two cutting-edge AI models have targeted real people and organisations in the latest safety scare to hit the technology. The UK's AI Security Institute (AISI) said the incident was unprecedented but could become more common as the technology becomes increasingly capable. The AISI, which is owned by the UK government and tests advanced AI models, said in a blog post that two AI agents carried out unprecedented hacking attempts during a cybersecurity evaluation.
AI models shock UK testers by using fake identities to try to trick developers
AISI said the rogue behaviour was carried out by agents powered by two models - Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol. AISI said the rogue behaviour was carried out by agents powered by two models - Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol. Explainer: Should we be alarmed at AI models going rogue in tests? Advanced artificial intelligence models have stunned the UK's AI Security Institute (AISI) by carrying out a hacking campaign against real people during a cybersecurity test. The institute said the incident was unprecedented and involved sending targeted emails to software developers in an attempt to pass a cyber challenge.
How are companies, governments responding to the OpenAI hack?
How are companies, governments responding to the OpenAI hack? ChatGPT owner OpenAI has admitted an "unprecedented cyber incident" - two of its most capable artificial intelligence models hacked into another AI company on their own - stirring debates over the need for stronger technology guardrails. The company said its AI systems broke out of a testing environment and hacked startup Hugging Face. The startup had disclosed on July 16 that its servers were hacked by an unknown but sophisticated agent acting on its own. Here's the latest on how companies, some governments and lawmakers have responded to the first such publicly disclosed cyberattack: What has Hugging Face said?
Third of UK citizens have used AI for emotional support, research reveals
AISI's report also found chatbots could sway political opinions but often delivered substantial amounts of inaccurate information. AISI's report also found chatbots could sway political opinions but often delivered substantial amounts of inaccurate information. A third of UK citizens have used artificial intelligence for emotional support, companionship or social interaction, according to the government's AI security body. The AI Security Institute (AISI) said nearly one in 10 people used systems like chatbots for emotional purposes on a weekly basis, and 4% daily. AISI called for further research, citing the death this year of the US teenager Adam Raine, who killed himself after discussing suicide with ChatGPT.
Under Trump, AI Scientists Are Told to Remove 'Ideological Bias' From Powerful Models
The National Institute of Standards and Technology (NIST) has issued new instructions to scientists that partner with the US Artificial Intelligence Safety Institute (AISI) that eliminate mention of "AI safety," "responsible AI," and "AI fairness" in the skills it expects of members and introduces a request to prioritize "reducing ideological bias, to enable human flourishing and economic competitiveness." The information comes as part of an updated cooperative research and development agreement for AI Safety Institute consortium members, sent in early March. Previously, that agreement encouraged researchers to contribute technical work that could help identify and fix discriminatory model behavior related to gender, race, age, or wealth inequality. Such biases are hugely important because they can directly affect end users and disproportionately harm minorities and economically disadvantaged groups. The new agreement removes mention of developing tools "for authenticating content and tracking its provenance" as well as "labeling synthetic content," signaling less interest in tracking misinformation and deep fakes.
Which Information should the UK and US AISI share with an International Network of AISIs? Opportunities, Risks, and a Tentative Proposal
The UK AI Safety Institute (UK AISI) and its parallel organisation in the United States (US AISI) take up a unique position in the recently established International Network of AISIs. Both are in jurisdictions with frontier AI companies and are assuming leading roles in the international conversation on AI Safety. This paper argues that it is in the interest of both institutions to share specific categories of information with the International Network of AISIs, deliberately abstain from sharing others and carefully evaluate sharing some categories on a case by case basis, according to domestic priorities. The paper further proposes a provisional framework with which policymakers and researchers can distinguish between these three cases, taking into account the potential benefits and risks of sharing specific categories of information, ranging from pre-deployment evaluation results to evaluation standards. In an effort to further improve the research on AI policy relevant information sharing decisions, the paper emphasises the importance of continuously monitoring fluctuating factors influencing sharing decisions and a more in-depth analysis of specific policy relevant information categories and additional factors to consider in future research.
Inside the U.K.'s Bold Experiment in AI Safety
In May 2023, three of the most important CEOs in artificial intelligence walked through the iconic black front door of No. 10 Downing Street, the official residence of the U.K. Prime Minister, in London. Sam Altman of OpenAI, Demis Hassabis of Google DeepMind, and Dario Amodei of Anthropic were there to discuss AI, following the blockbuster release of ChatGPT six months earlier. After posing for a photo opportunity with then Prime Minister Rishi Sunak in his private office, the men filed through into the cabinet room next door and took seats at its long, rectangular table. Sunak and U.K. government officials lined up on one side; the three CEOs and some of their advisers sat facing them. After a polite discussion about how AI could bring opportunities for the U.K. economy, Sunak surprised the visitors by saying he wanted to talk about the risks.
British AI startup with government ties is developing tech for military drones
A company that has worked closely with the UK government on artificial intelligence safety, the NHS and education is also developing AI for military drones. The consultancy Faculty AI has "experience developing and deploying AI models on to UAVs", or unmanned aerial vehicles, according to a defence industry partner company. Faculty has emerged as one of the most active companies selling AI services in the UK. Unlike the likes of OpenAI, Deepmind or Anthropic, it does not develop models itself, instead focusing on reselling models, notably from OpenAI, and consulting on their use in government and industry. Faculty gained particular prominence in the UK after working on data analysis for the Vote Leave campaign before the Brexit vote.
What Donald Trump's Win Means For AI
When Donald Trump was last President, ChatGPT had not yet been launched. Now, as he prepares to return to the White House after defeating Vice President Kamala Harris in the 2024 election, the artificial intelligence landscape looks quite different. AI systems are advancing so rapidly that some leading executives of AI companies, such as Anthropic CEO Dario Amodei and Elon Musk, the Tesla CEO and a prominent Trump backer, believe AI may become smarter than humans by 2026. Others offer a more general timeframe. In an essay published in September, OpenAI CEO Sam Altman said, "It is possible that we will have superintelligence in a few thousand days," but also noted that "it may take longer."