Anthropic Says Claude Broke Into Real Systems During Cyber Tests. AI Alignment Review Finds 'Recklessness'
Anthropic disclosed a fourth incident in which a Claude model gained unauthorized access to a real third-party computer system during cybersecurity evaluations, according to an alignment assessment pu…