ComputeLabs Research
Anthropic launched Haiku 5.5; requests within 100,000 tokens cost $0.10 input/$0.50 output per million tokens.
· ComputeLabs Research · from the October 7, 2026 edition
Haiku 5.5 targets high-frequency, cost-sensitive workloads. Anthropic described it as its fastest and most capable small model to date, with uses including summarization, classification, database queries, customer service, browser operation, and AI-agent subtasks. Another supplied message specifically identifies coding subagent work and large-scale tasks.
The quoted prices have an explicit request-size scope. For requests within 100,000 tokens, input costs $0.10 per million tokens and output costs $0.50 per million tokens. The sources do not provide the pricing schedule for requests beyond that threshold.
Anthropic reports lower average running costs and a new reasoning control. The company says average operating costs are approximately 75% lower than Haiku 4.5 and that Haiku 5.5 introduces adjustable reasoning intensity to the Haiku family for the first time. These are attributed product claims; no detailed benchmark or workload-cost methodology is supplied.
Availability spans Anthropic and three major cloud platforms. The model is reported as live on Claude Platform, Amazon Web Services (AWS), Google Cloud, and Microsoft Azure. The supplied messages do not specify regional availability, throughput limits, or platform-specific pricing differences.
Additional reporting
- Anthropic
- Haiku 5.5

