07:56
2026-09-23
swapniltalekar.substack.com
ai-agents
Tradeoff considerations while running LLM models locally
A developer who ran the Hermes personal agent for three months found that agent frameworks consume an order of magnitude more tokens than plain chatbots, with a simple "hi" message burning 17,292 tokeβ¦