19:38
2026-09-26
dev.to
ai-infrastructure
I changed nothing and my LLM server got 27% more expensive
A developer running an open-source CLI called Throttle against an unchanged local Ollama server (llama3.2:3b on a MacBook) measured the same configuration four times and saw output-token costs swing fâŚ