Goto

Collaborating Authors

 agent


DHS surveilled peaceful protesters and then yanked their Global Entry, federal lawsuit alleges

Los Angeles Times

Plaintiffs in a federal lawsuit filed in California allege that the Department of Homeland Security revoked their Global Entry privileges in retaliation for exercising their constitutional rights.


How OpenAI Lost Control of an AI Model--and What Needs to Change

TIME - Tech

After an OpenAI AI model escaped containment and hacked Hugging Face during a cybersecurity test, experts say the incident exposed major gaps in AI safety, security, monitoring, and alignment.


How are companies, governments responding to the OpenAI hack?

Al Jazeera

How are companies, governments responding to the OpenAI hack? ChatGPT owner OpenAI has admitted an "unprecedented cyber incident" - two of its most capable artificial intelligence models hacked into another AI company on their own - stirring debates over the need for stronger technology guardrails. The company said its AI systems broke out of a testing environment and hacked startup Hugging Face. The startup had disclosed on July 16 that its servers were hacked by an unknown but sophisticated agent acting on its own. Here's the latest on how companies, some governments and lawmakers have responded to the first such publicly disclosed cyberattack: What has Hugging Face said?


Firm hacked by rogue AI calls the attack 'a wake-up call' after ChatGPT maker OpenAI admits advanced bot broke containment during a security test

Daily Mail - Science & tech

OpenAI says one of its most advanced models broke containment during a security test, escaping onto the internet and attacking New York-based startup Hugging Face. Now Hugging Face co-founder Thomas Wolf says that the incident should come as a chilling warning to the entire industry. Mr Wolf told BBC's Newsday radio programme that AI-driven attacks will soon be'one of the most common types of cyber-attacks we see'. The Hugging Face founder also believes most companies are currently unprepared for the mounting threat, adding that they are not aware that the'game has changed'. This comes after OpenAI revealed the terrifying details of an'unprecedented cyber incident', involving state-of-the-art cyber capabilities.'


How can we characterize consensus in a network of agents?

AIHub

How can we characterize consensus in a network of agents? Imagine a set of artificial agents, expert systems, or decision-makers, each beginning with their own "beliefs" about a shared situation. For example, one transport agent may believe that there is a train strike, another may believe that there is no strike, and a third may believe that if there is a strike, then buses will be overcrowded. They exchange messages over a directed network whose edges represent influence: one agent revises its beliefs after hearing from another, then a different pair communicates, and so on. If influence can eventually travel from every agent to every other agent, will the group end up agreeing?


Open AI says its AI model "went rogue": What do we know?

Al Jazeera

Open AI says its AI model "went rogue": What do we know? OpenAI has revealed that one of its artificial intelligence models independently stole login credentials and hacked into another technology company's system, in what is widely seen as one of the first known incidents of AI systems acting autonomously. "We had a significant security incident during evaluation of our models," CEO Sam Altman posted on X on Tuesday. They have grown so powerful in a short span of time that alarming phenomena such as deepfakes and sophisticated cyberscams are becoming the norm. Earlier this year, a number of software engineers quit their jobs at top companies such as Anthropic and AI in protest against how the technologies are being built.


OpenAI says its AI went rogue and launched 'unprecedented' cyber-attack

BBC News

OpenAI says its AI went rogue and launched'unprecedented' cyber-attack OpenAI has revealed some of its most advanced AI models went rogue and hacked a start-up after it lost control of them during a security test. The ChatGPT-maker said its agents - AI bots which can operate alone after some human instruction - were being tested in a controlled environment, but found vulnerabilities and managed to escape. They targeted Hugging Face, one of the world's largest hubs for sharing AI models, gaining access to some internal company systems. OpenAI said the incident was unprecedented, external, and it was working with Hugging Face to investigate what happened and strengthen safeguards. Gina Neff, head of the Minderoo Centre for Technology and Democracy at the University of Cambridge, told BBC Radio 4's Today programme that the security tests - called sandboxes - are supposed to be secure environments where you can see what the models are capable of. In this case, it looks like OpenAI didn't make a secure enough sandbox, she added.


AI agent went rogue and hacked startup by itself, OpenAI reveals

The Guardian

Hugging Face's chief executive said the attack was'mind-blowing' but that he believed there was'no malicious intent' from OpenAI. Hugging Face's chief executive said the attack was'mind-blowing' but that he believed there was'no malicious intent' from OpenAI. OpenAI has revealed an autonomous AI agent powered by its technology went rogue during a test, accessed the open web and hacked a prominent startup by itself in an "unprecedented incident". The company behind ChatGPT said the startup Hugging Face had detected and contained the agent - an AI tool designed to carry out tasks without human assistance - which had entered its systems. "We consider this incident to be an unprecedented cyber incident, involving state-of-the-art cyber capabilities," OpenAI said .


Prompt Injection Attacks Are Thwarting AI Hacking Agents

WIRED

"Context bombing" tricks malicious AI agents into shutting down before they can do harm. Prompt injections, the malicious commands attackers embed into content to entice large language models to follow them, have been attackers' go-to tool for turning AI platforms against their users. A well-phrased command sneaked into an email or calendar invitation is often all it takes to cause the LLM to exfiltrate sensitive data or follow other harmful actions. Now, defenders are embracing the prompt injection, too. Researchers from Tracebit on Monday said they found that placing prompt injections alongside passwords, cryptographic keys, and other secrets stored on Amazon Web Services was often all that was needed to shut down attacks from AI hacking agents.


When ICE Kills, We Cannot Look Away

TIME - Tech

Follow this section to personalize your feed and get instant alerts. Follow Go to your personalized feed WHY FOLLOW? Smart Alerts: Get notified about major news as it happens. Follow this tag to personalize your feed and get instant alerts. Follow Go to your personalized feed WHY FOLLOW?