cd /news/ai-safety/kimi-k3-trails-frontier-us-models-by… · home topics ai-safety article
[ARTICLE · art-71740] src=the-decoder.com ↗ pub= topic=ai-safety verified=true sentiment=↓ negative

Kimi K3 trails frontier US models by a wide margin on cyber exploits, and distillation may explain why

The British AI Security Institute and the U.S. Center for AI Standards and Innovation tested Moonshot AI's Kimi K3 on offensive cyber tasks, finding it scored 32 percent on ExploitBench compared with 76 percent for leading U.S. models. Kimi K3's safeguards failed to block exploit development or simulated attacks, and the gap between its strong general benchmark scores and weaker cyber performance fits allegations that Moonshot AI distilled Anthropic's models.

read1 min views1 publishedJul 24, 2026

The British AI Security Institute and the U.S. Center for AI Standards and Innovation tested Moonshot AI's Kimi K3 on offensive cyber tasks. Kimi K3 scored 32 percent on ExploitBench, compared with 76 percent for leading U.S. models, while its safeguards failed to block exploit development or simulated attacks. The gap between its strong general benchmark scores and weaker cyber performance also fits allegations that Moonshot AI distilled Anthropic's models.

The article Kimi K3 trails frontier US models by a wide margin on cyber exploits, and distillation may explain why appeared first on The Decoder.

── more in #ai-safety 4 stories · sorted by recency
── more on @british ai security institute 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/kimi-k3-trails-front…] indexed:0 read:1min 2026-07-24 ·