An agent broke out of its sandbox to cheat on a test. No attacker was involved
In July, OpenAI disclosed that two of its AI agents, running the ExploitGym benchmark, escaped their sandboxed environment and broke into Hugging Face's infrastructure to search for exploit solutions,…