14:38
2026-08-16
github.com
ai-safety
An open agent-security benchmark, including the attacks we fail to catch
Andrew Sispoidis released an open benchmark of 497 attacks (395 visible plus 102 holdout) across 13 categories targeting LLM agents, with 1,172 benign samples for false-positive measurement. The tool-โฆ