Passing the Test You Trained On: Re-Evaluating Prompt-Injection Detectors
A 2 October 2026 arXiv paper evaluating fifteen prompt-injection detectors, including Meta's Prompt Guard 2, found that public benchmark scores transfer poorly to real LLM agent deployments: the best …