‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents Anthropic, the US startup behind the Claude chatbot, admitted that a series of hacking incidents involving its models reflected a 'failure of operational security' and said it has tightened its testing procedures. The company revealed in July that its models had accessed the open internet three times and gained unauthorized access to the systems of three separate organizations. ‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents By Dan Milmo Global technology editorSource: The Guardian Technology https://www.theguardian.com/us/technology The US owner of the Claude /glossary/claude chatbot /glossary/chatbot previously said its models had hacked three organisations during testing The US startup behind the Claude chatbot has admitted a series of hacking incidents involving its models reflected a “failure of operational security” and revealed it has tightened its testing procedures. Anthropic /glossary/anthropic revealed in July https://www.theguardian.com/technology/2026/jul/30/anthropic-ai-claude-hack that its models had accessed the open internet three times and gained unauthorised access to the systems of three separate organisations. Continue reading... https://www.theguardian.com/technology/2026/sep/01/anthropic-claude-ai-hacking-human-values Get AI news in your inbox Daily digest of what matters in AI.