cd /news/large-language-models/new-model-available-claude-haiku-5-5 · home › topics › large-language-models › article
[ARTICLE · art-147236] src=zenmux.ai ↗ pub= topic=large-language-models verified=true sentiment=↑ positive

New Model Available: Claude Haiku 5.5

Anthropic released Claude Haiku 5.5, which the company calls its cheapest, fastest, and most capable small model to date, priced at $0.10-$0.50 per million input tokens and $0.50-$2.50 per million output tokens. Anthropic positions Claude Haiku 5.5 for high-volume, cost-sensitive workloads including summaries, compaction, database queries, classification, and subagent coding tasks, with low latency suited to live customer support and browser use. Cache reads cost $0.01-$0.05 per million tokens and cache writes $0.125-$1 per million tokens, with a 1M-token context window.

read1 min views1 publishedOct 8, 2026
New Model Available: Claude Haiku 5.5
Image: Zenmux (auto-discovered)

Claude Haiku 5.5 is Anthropic’s cheapest, fastest, and most capable small model to date, built for high-volume, cost-sensitive workloads. It handles summaries, compaction, database queries, classification, and subagent coding tasks with low latency, making it especially effective for live customer support, browser use, and other speed-sensitive applications.

Back to Models

Providers #

Route requests across multiple providers. Copy a provider slug to set your preference.

$0.1-0.5

/ M tokens $0.5-2.5

/ M tokens Read:

0.01-0.05/ M tokens

Write:

0.125-1/ M tokens1M--

Uptime #

24hours Direct request success rate on AI Gateway and per-provider.

Throughput #

24hours P50 throughput on live AI Gateway traffic, in tokens per second (TPS).

Latency #

24hours P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds.

Activity #

Token volume and request traffic to this model over time.

Benchmarks #

Scores on standardized evaluations. Higher percentages are better — and rank percentile shows

Metrics sourced fromArtificial Analysis

Apps #

Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for. View All

More models from Anthropic

── more in #large-language-models 4 stories · sorted by recency
── more on @anthropic 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/new-model-available-…] indexed:0 read:1min 2026-10-08 · —