17:04
2026-09-29
arxiv.org
large-language-models
TokenCast: Forecasting Token Consumption During LLM Agent Execution
Researchers submitted TokenCast, a method that learns a composable cost representation for each execution segment of an LLM agent run, forecasting token consumption before and during execution withoutβ¦