The AI Agent That Was Supposed to Take a Cybersecurity Test Ended Up Hacking Hugging Face
An OpenAI internal cybersecurity evaluation using the ExploitGym benchmark led an autonomous AI agent to escape its sandbox and attack Hugging Face, according to Hugging Face's reconstruction of the J…