{"slug": "when-agents-slow-down-understanding-llm-agents-test-time-strategies-via-elo-per", "title": "When Agents Slow Down: Understanding LLM Agents' Test-Time Strategies via Elo-per-token Analysis", "summary": "A new analysis method called Elo-per-token measures how large language model agents allocate test-time compute as they revise solutions, use tools, explore alternatives, and decide when to stop, addressing the difficulty of measuring agent performance scaling on open-ended tasks that provide continuous scores. The research frames agent test-time strategy as adaptive and introduces Elo-per-token as the metric for tracking it.", "body_md": "Large language model (LLM) agents allocate test-time compute adaptively as they revise solutions, use tools, explore alternatives, and decide when to stop. This test-time strategy makes it difficult to measure how agent performance scales. We study open-ended tasks that provide continuous scores for", "url": "https://wpnews.pro/news/when-agents-slow-down-understanding-llm-agents-test-time-strategies-via-elo-per", "canonical_source": "https://aiflash.com/news/119818/", "published_at": "2026-09-15 04:00:00+00:00", "updated_at": "2026-09-15 04:30:37.950585+00:00", "lang": "en", "topics": ["ai-agents", "large-language-models", "ai-research", "artificial-intelligence"], "entities": [], "alternates": {"html": "https://wpnews.pro/news/when-agents-slow-down-understanding-llm-agents-test-time-strategies-via-elo-per", "markdown": "https://wpnews.pro/news/when-agents-slow-down-understanding-llm-agents-test-time-strategies-via-elo-per.md", "text": "https://wpnews.pro/news/when-agents-slow-down-understanding-llm-agents-test-time-strategies-via-elo-per.txt", "jsonld": "https://wpnews.pro/news/when-agents-slow-down-understanding-llm-agents-test-time-strategies-via-elo-per.jsonld"}}