22:20
2026-09-29
simonwillison.net
ai-safety
Quoting Anthropic Frontier Red Team
Anthropic's Frontier Red Team reported that GLM-5.3 achieved full control flow hijacks in 4% of 100 randomly selected tasks from its internal Binary Exploitation benchmark, while Claude Mythos Previewβ¦