Prompt Caching, Batches API, and Model Routing to Cut LLM Costs
Anthropic's prompt caching, Batches API, and model routing can significantly reduce LLM costs without switching providers. Prompt caching reuses prefixes at 0.1× the base input price, the Batches API …