LLM agents often lack the operational knowledge to act reliably in new environments, as they must discover specific tool behaviors or environment conventions on their own. Without memory of past attempts, they repeat the same mistakes across tasks, leading to more task failures and longer trajectori
Wikipedia's Rogue Agent Forensics: What Wikimedia's Investigation Reveals About Uncontrolled Agent Swarms