14:03
2026-09-17
arxiv.org
ai-safety
AgentLSD: Evaluating AI Security Agents Under Adversarial Task Contamination
Researchers submitted AgentLSD, a controlled framework for evaluating AI security agents under adversarial task contamination, to arXiv on 16 Sep 2026. Testing six models on 11 web Capture the Flag chβ¦