02:07
2026-08-29
arxiv.org
ai-safety
Safety Does Not Compose: Non-Decaying Loop State for Autonomous LLM Agents
A new arXiv paper (2608.27141) demonstrates that safety monitors for autonomous large language model (LLM) agents fail to detect attacks whose evidence is spread across multiple iterations, because tr…