Claude Haiku 5.5

Claude Haiku 5.5 has been Anthropic’s fastest model in the 5.5 family since early October 2026, built for high-volume tasks such as summarization, classification, browser use, and subagents. Context grows from 200,000 to one million tokens, output to 128,000, and average runtime costs are around 75 percent below Haiku 4.5. New to the Haiku class: adaptive reasoning with effort control. Text and image as input, proprietary, API-only.

Anthropic Version 5.5 Commercial use permitted Dense 1000 K Context 06/2026 $0.1 / $0.5 per 1M

  • Proprietary
  • Frontier
  • Anthropic
  • Text
  • Vision
  • Instruction-Tuned
  • Real-Time

Sovereign Risk: MEDIUM The model is developed and operated by Anthropic, a US-based company. The ‘medium’ rating stems from the provider’s US jurisdiction. Laws such as the CLOUD Act could theoretically allow US authorities to access data processed on Anthropic’s servers. As this is a cloud-only model, this risk cannot be mitigated through local deployment. For users outside the US — particularly in jurisdictions with strict data protection requirements such as the GDPR — this represents a potential sovereignty risk.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
77.33
Routine
46.32
Reasoning
31.01

Rank #18

LLM Judge Avg
3.91
100 Coverage
Avg Task Duration
13.6
Real-Time
Token Rate
196.69
Output Rate
P95 Latency
42.99
Top 5 %
Total Tokens
189600
Output Volume
Cost per 1K
$0.0005
USD / 1K Requests
Benchmark Cost
$0.09
Total · 189600 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Claude Haiku 5.5 Best model Ø All models
Code Quality 84.92
CLI Benchmark 88.33
Logical Reasoning 76.15
UX Writing 72.71
Documentation 70.74
Content Transform. 79.58
Cultural Intelligence 74.52
Synthesis Quality 66.67
Tool Execution 90
ToolUse Score 77.17
Benchmark Cost $0.09

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile