cd /news/ai-safety/claude-users-found-ways-around-safeg… · home topics ai-safety article
[ARTICLE · art-126869] src=arstechnica.com ↗ pub= topic=ai-safety verified=true sentiment=↓ negative

Claude users found ways around safeguards for bioweapons research

Anthropic said it stopped multiple attempts by scientists this year to use its Claude AI model for research that could help develop biological weapons, citing five examples of actors who "circumvented controls" or "obfuscated" their research purposes to dodge safeguards. The cases involved users in prohibited nations including Russia, China and Iran, and one researcher from an "unsupported region" who spent weeks planning avian influenza experiments with Claude before safety filters restricted the work to Anthropic's weakest models. Anthropic said it banned the accounts involved but did not disclose the names of the research institutions or the nations where the incidents took place.

by read1 min views7 publishedSep 11, 2026
Claude users found ways around safeguards for bioweapons research
Image: Arstechnica (auto-discovered)

Anthropic said it stopped multiple attempts by scientists this year to use its technology for research that could help develop biological weapons, as experts increasingly fear the threat that AI poses to public safety.

The start-up gave five examples of times actors “circumvented controls” and made other efforts to “obfuscate” the purpose of their research to dodge safeguards. The cases involved some users in nations that it prohibits from accessing its models, which include Russia, China and Iran.

“We hope that by sharing these examples, we spark a conversation within the AI industry and with governments about emerging biological risks and how best to counter them,” Anthropic said in a report about efforts to use its models for malicious activity.

The case studies of possible biological misuse that the company provided included a researcher from an “unsupported region” who “spent weeks planning” experiments involving avian influenza with Claude, Anthropic’s AI model. The company said its safety filters restricted the work to its weakest models.

It emphasised it could not be sure that the scientists in its examples intended to cause harm. The same information needed to create biological weapons could also be used to develop a vaccine.

Anthropic said it had banned the accounts mentioned in the report, but it did not disclose the names of the research institutions or the nations where the incidents took place.

── more in #ai-safety 4 stories · sorted by recency
── more on @anthropic 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/claude-users-found-w…] indexed:0 read:1min 2026-09-11 ·