# Implausible generation speed number in Assistant

> Source: <https://kagifeedback.org/d/11428-implausible-generation-speed-number-in-assistant/1>
> Published: 2026-09-08 00:28:01+00:00

Please see [this Assistant thread](https://assistant.kagi.com/share/64df0512-c47b-4c1e-8832-62fa98300521) as an example.

48,938 tokens in 197 seconds works out to 248 tokens per second, yet the UI shows 22 tokens per second, so this is confusingly off by an order of magnitude.

(It might be plausible that the 22 tok/s measures only the final output tokens, and disregards reasoning and tool call/response tokens. But it would be a little weird to measure this way over the total request duration, because the LLM is spending most of its time reasoning and doing tool calls.)

I would expect to see a number around 248.
