New Model Available: Claude Haiku 5.5 Anthropic released Claude Haiku 5.5, which the company calls its cheapest, fastest, and most capable small model to date, priced at $0.10-$0.50 per million input tokens and $0.50-$2.50 per million output tokens. Anthropic positions Claude Haiku 5.5 for high-volume, cost-sensitive workloads including summaries, compaction, database queries, classification, and subagent coding tasks, with low latency suited to live customer support and browser use. Cache reads cost $0.01-$0.05 per million tokens and cache writes $0.125-$1 per million tokens, with a 1M-token context window. Claude Haiku 5.5 is Anthropic’s cheapest, fastest, and most capable small model to date, built for high-volume, cost-sensitive workloads. It handles summaries, compaction, database queries, classification, and subagent coding tasks with low latency, making it especially effective for live customer support, browser use, and other speed-sensitive applications. Back to Models https://zenmux.ai/models Providers Route requests across multiple providers. Copy a provider slug to set your preference. $0.1-0.5 / M tokens $0.5-2.5 / M tokens Read: 0.01-0.05 / M tokens Write: 0.125-1 / M tokens1M-- Uptime 24hours Direct request success rate on AI Gateway and per-provider. Throughput 24hours P50 throughput on live AI Gateway traffic, in tokens per second TPS . Latency 24hours P50 time to first token TTFT on live AI Gateway traffic, in milliseconds. Activity Token volume and request traffic to this model over time. Benchmarks Scores on standardized evaluations. Higher percentages are better — and rank percentile shows Metrics sourced from Artificial Analysis https://artificialanalysis.ai/ Apps Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for. View All https://zenmux.ai/analytics/apps Related Models More models from Anthropic https://zenmux.ai/anthropic