From LLM Inference to Agentic Workloads: Characterization and Implications
A new benchmark suite, AgentSysBench, reveals that agentic AI workloads differ fundamentally from conventional LLM serving, with non-LLM components dominating latency in 5 of 10 applications and sandbox memory peaking at…