{"slug": "redis-built-a-cache-that-cuts-llm-costs", "title": "Redis built a cache that cuts LLM costs", "summary": "Redis's LangCache delivers a 70% cache hit rate that cuts LLM spend by 70% and runs 4X faster for a voice app used in patient care, according to a customer quoted by Redis. The customer said the app handles specific treatment questions and requires absolute accuracy, and credited LangCache with addressing cost concerns for high usage while improving real-time patient interactions.", "body_md": "\"Our voice app for patient care gets a lot of specific treatment questions, so it has to be absolutely accurate, and that's what LangCache does. I was worried about LLM costs for high usage, but with LangCache, we're getting a 70% cache hit rate, which saves 70% of our LLM spend. On top of that, it’s 4X faster, which makes a huge difference for real-time patient interactions.\"", "url": "https://wpnews.pro/news/redis-built-a-cache-that-cuts-llm-costs", "canonical_source": "https://redis.io/langcache/", "published_at": "2026-09-11 08:54:21+00:00", "updated_at": "2026-09-11 09:01:50.596173+00:00", "lang": "en", "topics": ["ai-infrastructure", "large-language-models", "ai-products"], "entities": ["Redis", "LangCache"], "alternates": {"html": "https://wpnews.pro/news/redis-built-a-cache-that-cuts-llm-costs", "markdown": "https://wpnews.pro/news/redis-built-a-cache-that-cuts-llm-costs.md", "text": "https://wpnews.pro/news/redis-built-a-cache-that-cuts-llm-costs.txt", "jsonld": "https://wpnews.pro/news/redis-built-a-cache-that-cuts-llm-costs.jsonld"}}