Implausible generation speed number in Assistant A Kagi Assistant user reports that the displayed generation speed of 22 tokens per second contradicts the actual rate of 248 tokens per second calculated from 48,938 tokens generated in 197 seconds, suggesting the UI may be undercounting by an order of magnitude. The user speculates that the metric might exclude reasoning and tool call tokens, but notes this would be misleading given the total request duration. Please see this Assistant thread https://assistant.kagi.com/share/64df0512-c47b-4c1e-8832-62fa98300521 as an example. 48,938 tokens in 197 seconds works out to 248 tokens per second, yet the UI shows 22 tokens per second, so this is confusingly off by an order of magnitude. It might be plausible that the 22 tok/s measures only the final output tokens, and disregards reasoning and tool call/response tokens. But it would be a little weird to measure this way over the total request duration, because the LLM is spending most of its time reasoning and doing tool calls. I would expect to see a number around 248.