Goto

Collaborating Authors

 guardrail


MAGA mom's dire warning over tech guardrails as Trump puts pedal to the metal on race with China

FOX News

Humans First chairwoman Amy Kremer calls for AI safety guardrails to protect children from chatbot dangers and prevent artificial intelligence from displacing American workers.


Chinese AI tool told researchers how to make bioweapons

BBC News

Chinese AI developer Moonshot is conducting an internal review after researchers were able to persuade two of its popular Kimi models to tell them how to make biological weapons and carry out assassinations. Mindgard, which tests the security of AI systems, told the BBC it discovered in July that Kimi K2.6 and K3 Swarm could evade safety limits put in place by developers. It arose during a process called jailbreaking, where researchers use a series of complex instructions to see if AI tools ignore guardrails - which Mindgard said should have stopped Kimi from discussing concerning topics. Moonshot told the BBC it welcomed third-party input as a key pillar for building better and safer AI. The company also told the BBC it was in discussion with Mindgard about its findings.


What's the US–China AI 'hotline' that Trump plans to pitch to Xi Jinping?

Al Jazeera

What's the US-China AI'hotline' that Trump plans to pitch to Xi Jinping? Share What's the US-China AI'hotline' that Trump plans to pitch to Xi Jinping? on social media The United States wants a new "notification mechanism" with China to warn each other when an artificial intelligence incident becomes serious enough to threaten national security. The proposal, effectively an "AI hotline", was discussed on Sunday during talks in New York between US Treasury Secretary Scott Bessent and Chinese Vice Premier He Lifeng, ahead of a summit between US President Donald Trump and Chinese President Xi Jinping in Washington this week. The initiative comes amid an intensifying contest for AI supremacy between the world's top two economies, together with increasing warnings from tech giants that guardrails are necessary around the ever-growing technology. So, what do we know about this so-called hotline, and where is the AI battle between the US and China headed?


How, Exactly, Could A.I. Kill Us?

The New Yorker

How, Exactly, Could A.I. Kill Us? Employees of A.I. companies are increasingly sounding the alarm. When Jacob Coxon, a mathematician and software engineer, resigned from his research job at Anthropic last week, he warned, "The people building AI earnestly believe that it could kill us all by the end of the decade." This would be a remarkable statement were it not for the fact that artificial-intelligence leaders have long been saying precisely this. Dario Amodei, then a research scientist at OpenAI, raised the concern that a superintelligence "could destroy humanity," adding, "I can't see any reason and principle why that couldn't happen." Earlier, in 2015, Sam Altman, just before he co-founded OpenAI, said, "I think A.I. will probably most likely lead to the end of the world, but in the meantime, there'll be great companies created with serious machine learning." Elon Musk, in 2014: "I think we should be very careful about artificial intelligence. If I were to guess at what our biggest existential threat is, it's probably that." Perhaps the only thing that's changed between then and now is that the rest of the world is finally paying attention. In recent weeks, the same technology that, a couple of years ago, couldn't count the number of "R"s in the word "strawberry"--and, a couple of days ago, insisted to me that Dolly Parton is still alive--has been used to solve the Navier-Stokes problem, which has been stumping mathematicians for nearly a century, and has also demonstrated its ability to go rogue in a series of disturbing hacking incidents. Last week, Anthropic also published a report detailing various ways in which bad actors have attempted to use the company's A.I. models, including one especially troubling case of a scientist using Claude to study a virus at a military research institute--work that could yield a vaccine, a biological weapon, or both. A few days later, Amodei published a letter calling for an industry-wide slowdown and more government regulation, to which President Donald Trump responded, on Truth Social, "The only control or'guardrails' that AI needs is a STRONG AND SMART (High IQ!) PRESIDENT, and the U.S.A. has that, in spades!" I recently spoke on The Political Scene podcast with my colleague Joshua Rothman, a staff writer who has been covering A.I. for years, about whether we're all doomed, and what it would even look like for A.I. to destroy humanity. Can A.I. leaders save us from their own creation, and how can the government coöperate in order to do so? And is A.I.'s capacity to do good--its potential to mitigate climate change or innovate medical treatments--hopelessly intertwined with its capacity to do bad? Our conversation has been edited for length and clarity. A lot of people in the world of artificial intelligence are talking about their P(doom) number, which is the probability that artificial intelligence will lead to an absolutely catastrophic situation--possibly, or probably, killing us all.


US speaker rejects AI pause, warns of losing edge to China

Al Jazeera

US House Speaker Mike Johnson has rejected calls to pause artificial intelligence development, arguing that doing so could give China an advantage over the United States. "We cannot have a moratorium on the development of AI," Johnson told reporters on Tuesday. "Because then we will lose our edge to China, and that has serious national security implications for every American family." They don't need the government to tell them to slow it down. If they want to slow it down, they should," he said. His comments come as calls to rein in AI are growing across the political spectrum amid concerns about what the technology could mean for jobs, energy costs and public safety. Unlikely allies US Senator Bernie Sanders, a self-described democratic socialist, and Steve Bannon, a former adviser to President Donald Trump and prominent voice on the US right, both called for restrictions on AI on Tuesday, as they appeared at the same event in Washington. Sanders warned of what he called the ...


Trump dismisses AI safety alarm, says U.S. already has tools to police industry

The Japan Times

Trump dismisses AI safety alarm, says U.S. already has tools to police industry U.S. President Donald Trump speaks during a meeting in the Oval Office at the White House in Washington on Sept. 2. | REUTERS WASHINGTON - President Donald Trump said on Monday that the U.S. already has guardrails in place to regulate and prosecute AI companies, appearing to play down concerns expressed by industry leaders over the weekend about the misuse of artificial intelligence harming people. Public alarm about the potential danger posed by AI is growing, after Anthropic researcher Jacob Coxon said last week he had resigned, in part because the "people building AI earnestly believe that it could kill us all by the end of the decade." AI-related stocks fell worldwide on Monday after leaders of some of the industry's largest companies called for slowing the pace of the technology's development. "The only control or'guardrails' that AI needs is a STRONG AND SMART (High IQ!) PRESIDENT, and the U.S.A. has that, in spades!" Trump wrote on Truth Social, adding that China would benefit from the doubt being cast on AI development. Trump is scheduled to meet with Chinese President Xi Jinping next week in Washington. In a time of both misinformation and too much information, quality journalism is more crucial than ever.


Trump claims he is only 'guardrail' needed to control AI as top Republicans join him in dismissing calls for more checks – live

The Guardian

Trump claims he is only'guardrail' needed to control AI as top Republicans join him in dismissing calls for more checks - live President condemns what he claims is a'sick conspiracy' amid global tech selloff following CEOs calls for slowing pace of development'I don't think Americans should be scared of anything': Vance tries to quell concerns about AI regulation Trump claims'sick conspiracy' over calls for more guardrails on AI companies Johnson: Congress shouldn't be in charge of AI guardrails, insists its tech companies responsibilities Johnson to discuss Trump's $5k midterm pledge, says incentive needs approval from Congress Trump's mail-in voting restrictions blocked by a second judge Trump claims'sick conspiracy' over calls for more guardrails on AI companies Donald Trump has claimed there is a "SICK conspiracy" "The only control or'guardrails' that AI needs is a STRONG AND SMART (High IQ!) PRESIDENT, and the U.S.A. has that, in spades!" Trump wrote on Truth Social, insisting that the administration has stopped leaders of AI companies from "doing bad, or potentially bad" However, he didn't point to any concrete examples for how he has curbed possible abuse or malfeasance in the industry. The president said over the weekend that the AI arms race is of the upmost importance to the US national security. This comes after company whistleblowers claimed that the technology has the potential to "kill all of humanity" within the decade. Trump accused Amodei of "now pretending to be a'perfect little angel'" There is a SICK conspiracy going on against AI and Data Centers, and the only one that is happy about it is China We are leading China, and all others, and will continue to do so. Conspiracy Theorists, Treasonists, Traitors, and Leakers, BEWARE! 'I don't think Americans should be scared of anything': Vance tries to quell concerns about AI regulation Trump claims'sick conspiracy' over calls for more guardrails on AI companies Johnson: Congress shouldn't be in charge of AI guardrails, insists its tech companies responsibilities Johnson to discuss Trump's $5k midterm pledge, says incentive needs approval from Congress Trump's mail-in voting restrictions blocked by a second judge Federal law already requires those seeking to become US permanent residents to prove they will not be a burden to the US - also known as a "public charge" - but the new rule includes broader range of programs that could disqualify them.


Trump dismisses calls for AI slowdown from leading tech CEOs

Al Jazeera

US President Donald Trump has downplayed the need for his administration to place checks on artificial intelligence development, saying he is worried about ceding the US's lead to China. Speaking to reporters during his trip to Ireland on Sunday, Trump also said he acknowledged the need for some regulation but did not provide further details on potential measures. "We can put guardrails, we can do this and that, but I think you have a lot of negative forces that are bringing it up that shouldn't be bringing it up, and they're bringing up things that won't happen," Trump told reporters after watching the Irish Open at his resort in Doonbeg. He said that with the US leading the development of AI over China, "frankly, I want to keep it that way because whoever wins AI wins." Trump's comments came a day after Anthropic CEO Dario Amodei published a blog post saying the AI industry must slow the pace of improving the technology's capabilities to allow time for guardrails.


I Let an AI Agent Hack All My Gadgets--and I'd Do It Again

WIRED

I Let an AI Agent Hack All My Gadgets--and I'd Do It Again After I removed the safety guardrails from a powerful open-source model, it found vulnerabilities in my household devices and hacked into a PC. But it also told me how to make everything a lot more secure. As the author of a newsletter about artificial intelligence, I consider it my duty to experience the bleeding edge of this technology firsthand. This week, that meant embracing some agentic mayhem. You're probably aware that frontier AI models have attained advanced cybersecurity capabilities in recent months.


OpenAI chief scientist warns no one is prepared for consequences of AI

BBC News

OpenAI's chief scientist Jakub Pachocki has called for extreme caution over AI's runaway progress and warned more intervention may be needed to ensure humans remain in control of the future. I am concerned no one is prepared for the consequences of a continued rapid rise in machine intelligence, he wrote in a blog post entitled An Alien Mind, external . The post comes only a few days after the ChatGPT-maker released its latest model GPT-6 Astra, which it called its most powerful product ever. OpenAI and other AI firms like Anthropic shared reports of their AI agents acting autonomously and carrying out real-world cyber-attacks on other companies. In July, OpenAI called an incident in which its AI agents - AI systems which can operate alone after human instruction - hacked the tech platform Hugging Face unprecedented.