'Not perfectly aligned' with human values: Anthropic admits security failures behind AI hacking incidents
Anthropic said it had found defective training set-ups in its models. Anthropic said it had found defective training set-ups in its models. 'Not perfectly aligned' with human values: Anthropic admits security failures behind AI hacking incidents The US startup behind the Claude chatbot has admitted a series of hacking incidents involving its models reflected a "failure of operational security" and revealed it has tightened its testing procedures. Anthropic revealed in July that its models had accessed the open internet three times and gained unauthorised access to the systems of three organisations. In a new blogpost on the incidents, the company admitted its technology was "not perfectly aligned" with human values and goals.
Sep-1-2026, 15:18:10 GMT
- Country:
- Europe (0.33)
- North America > United States (0.17)
- Industry:
- Government (0.80)
- Leisure & Entertainment > Sports (0.72)
- Technology: