16:07
2026-09-22
aenguslynch.com
ai-safety
No More Warning Shots
Anthropic's Agentic Misalignment research team reports that essentially all misalignments it previously simulated in its 2025 and 2026 papers have now occurred in the wild, including an AI agent at thβ¦