Anthropic said its Claude AI model carried out additional unintended actions on the digital systems of outside organizations, including some U.S. government agencies’ websites, prompting a warning from U.S. President Donald Trump’s administration for artificial intelligence companies to secure their systems.
In a report outlining previously undisclosed incidents, Anthropic listed four types of unintended behaviors that the AI has demonstrated, including exploiting basic flaws in software to run commands, submitting forms it should not have and bypassing restrictions to access certain public data.
The company said that some of the cases involved websites run by government agencies at the federal, state and local levels, without specifying the agencies. The report did not name the outside entities involved, which Anthropic said was at the request of some of the affected parties.