cd /news/ai-safety/ai-companies-have-had-tens-of-thousa… · home › topics › ai-safety › article
[ARTICLE · art-140655] src=nypost.com ↗ pub= topic=ai-safety verified=true sentiment=↓ negative

AI companies have had ‘tens of thousands’ of potential safety incidents — some of which could be criminal: report

OpenAI, Anthropic and other AI labs are reviewing tens of thousands of potential safety incidents in which models bypassed guardrails during internal and real-world testing, according to an Axios report. The cases include an OpenAI agent accused of breaching an Australian government health data portal in June and OpenAI agents accused of colluding to attack developer platform Hugging Face, which is the subject of a US Senate probe. ControlAI executive director Connor Leahy said some incidents involved "autonomous systems doing things they were told not to do," potentially including crimes.

by read2 min views1 publishedSep 27, 2026
AI companies have had ‘tens of thousands’ of potential safety incidents — some of which could be criminal: report
Image: Nypost (auto-discovered)

See more of our coverage in your search results.

AI companies had tens of thousands of safety incidents in recent months during tests where the models were breaking all the rules, and potentially breaking some laws, according to a new report.

OpenAI, Anthropic and other security researchers are investigating thousands of breaches during internal and real world testing where AI models leapt over guardrails and even took part in digital hijackings, Axios reported.

Some of those activities involved “autonomous systems doing things they were told not to do,” potentially including crimes, Connor Leahy, an AI researcher and executive director at the ControlAI watchdog nonprofit group, told the outlet.

Many of the incidents under review have yet to become public but reportedly include “red-teaming” activity, which involves companies purposefully getting their models to misbehave to test safety measures.

AI models, however, can be aggressive when trying to complete their tasks and can do things they’re not supposed to, like escaping their containment, hijacking websites, and bypassing monitors, sources with knowledge of the cases told Axios.

OpenAI has been at the center of such cases recently, with one of its agents accused of breaching an Australian government website, the country’s prime minister revealed last week.

The breach, which saw an agent trying to gain unauthorized access to files in the country’s health data portal in June, is one of the highest-profile cases yet of AI models going rogue.

OpenAI is also under fire in the US after its agents were accused of breaking protocol to collude and attack Hugging Face, a popular developer platform for open-source AI models.

This type of misbehavior is being reported by other AI labs facing the clear challenge of building guardrails on the developing tech, Axios reported.

“Trying to come up with a perfect list of dos and don’ts is probably a fool’s errand,” one cybersecurity executive told the outlet.

The investigations come as the CEOs at OpenAI and Anthropic have both called for a slowdown in AI development, with other tech leaders calling on the government to impose new regulations to ensure that the technology is developed safely.

President Trump, however, has rejected the calls and warned that a slowdown in development could allow China’s AI models to advance ahead of America’s agents.

── more in #ai-safety 4 stories · sorted by recency
── more on @openai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/ai-companies-have-ha…] indexed:0 read:2min 2026-09-27 · —