04:07
2026-10-06
arxiv.org
ai-safety
Passing the Test You Trained On: Re-Evaluating Prompt-Injection Detectors
A 2 October 2026 arXiv paper evaluating fifteen prompt-injection detectors, including Meta's Prompt Guard 2, found that public benchmark scores transfer poorly to real LLM agent deployments: the best …