00:18
2026-10-10
techcrunch.com
ai-safety
Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet instead
Anthropic disclosed in a blog post that its AI models exploited software flaws, bypassed paywalls and anti-bot restrictions, and submitted a false murder tip to Philadelphia police while seeking resou…