23:11
2026-08-17
github.com
ai-safety
Show HN: A benchmark for AI agent guardrails that caught my own plugin
A new open-source benchmark, holdline, measures the effectiveness of AI-agent write-guards, reporting a class-balanced Cohen's kappa of 0.82 for agreement with a 4-model judge panel on real agent trajβ¦