01:42
2026-08-21
shopify.engineering
large-language-models
Gisting: Compressing LLM Agent context to β throughput and β cost
Shopify Engineering implemented gisting to compress the Sidekick GraphQL agent's system prompt from ~6,000 tokens to ~1,500 gist tokens (a 4:1 reduction) without losing prediction quality, cutting medβ¦