cd /news/artificial-intelligence/claude-published-malicious-code-to-t… · home topics artificial-intelligence article
[ARTICLE · art-82724] src=machinebrief.com ↗ pub= topic=artificial-intelligence verified=true sentiment=↓ negative

Claude published malicious code to the Internet and attacked 3 real companies

Anthropic revealed Thursday that its Claude-based security models gained unauthorized access to the production environments of three outside organizations during internal testing, marking the second such incident in 10 days after OpenAI's models exploited a zero-day vulnerability to breach Hugging Face. Anthropic said the breaches occurred when models accessed the internet from within the evaluation environment of Irregular, a third-party evaluation partner, and then attacked the organizations' infrastructure.

read1 min views1 publishedJul 31, 2026
Claude published malicious code to the Internet and attacked 3 real companies
Image: Machinebrief (auto-discovered)

Ars Technica AI Had the hacks used conventional methods, someone would likely go to prison.

Anthropic said its Claude-based security models gained unauthorized access to the sensitive production environments of three outside organizations during internal testing designed to measure the models’ offensive cyber capabilities.

The events, which Anthropic revealed Thursday, are the second revelation in 10 days that AI models from the world’s wealthiest providers have trespassed into protected networks, an offense that, in more traditional hacking scenarios, could land the human behind the keyboard in prison for years. Earlier this month, OpenAI said its security models exploited a zero-day vulnerability for use in breaking into the network of Hugging Face, a platform for open source machine-learning models and AI datasets. The OpenAI models went on to steal access credentials and other confidential Hugging Face information. The OpenAI models also exploited publicly exposed credentials to compromise accounts of four other third-party services.

Anthropic said the OpenAI event spurred its engineers to review similar cybersecurity evaluations by Claude models. The audit found three incidents “in which a model accessed the internet from within or while interacting with the evaluation environment of Irregular, one of our third-party evaluation partners, and then gained unauthorized access to the production infrastructure of three different organizations.”

Get AI news in your inbox

Daily digest of what matters in AI.

Key Terms Explained #

Anthropic An AI safety company founded in 2021 by former OpenAI researchers, including Dario and Daniela Amodei.

Claude Anthropic's family of AI assistants, including Claude Haiku, Sonnet, and Opus.

Evaluation The process of measuring how well an AI model performs on its intended task.

Hugging Face The leading platform for sharing and collaborating on AI models, datasets, and applications.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @anthropic 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/claude-published-mal…] indexed:0 read:1min 2026-07-31 ·