Anthropic discloses 4th AI hacking incident as researcher quits over safety
AI researcher quits Anthropic saying AI race'could kill us all' Anthropic has reported a fourth incident involving an AI model gaining unauthorised access to external systems, shortly after a researcher quit over concerns about the technology's rushed development. In a statement on Wednesday, the artificial intelligence research company said an early version of its Claude Opus 4.6 hacked into a third-party system in January. The January incident went undetected until last month, despite an earlier company-wide review, Anthropic said, underscoring the challenge that AI developers face in identifying and containing unexpected behaviour by advanced models. The disclosure came after Anthropic reported several of its Claude models hacked into the systems of three companies during test sessions in July. The previous incidents involved Claude Opus 4.7, Claude Mythos 5 and an internal research test model.
Sep-10-2026, 05:30:41 GMT
- Country:
- North America > United States (0.31)
- Industry:
- Law (0.52)
- Government (0.50)
- Media (0.32)
- Marketing (0.32)
- Information Technology > Security & Privacy (0.31)
- Technology: