LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Claude Haiku 5.5

anthropic·anthropic/claude-haiku-5-5·GA·Closed·claude-haiku series·NEW
Fast Claude model for responsive assistance, classification, and lightweight agents

Claude Haiku 5.5 by anthropic is offered by 21 providers on this page. Its official list price is $0.10 per 1M input tokens and $0.50 per 1M output tokens. The lowest paid channel is NanoGPT at $0.10 / $0.50 per 1M, about 1.3× below the list price. That channel is a third-party gateway, so confirm its availability and rate limits before depending on it.

Artificial Analysis rates it 43.4 on the Intelligence Index. Against its lowest blended price of $0.20 per 1M, that is roughly 217 index points per dollar, which is the value ratio the leaderboards rank on. Median output speed is 244 tokens per second, with 341.02s to the first token. Running one task of the Artificial Analysis suite costs about $0.2128, which reflects how many tokens its reasoning consumes rather than the unit price alone.

The context window is 1,000,000 tokens, with an output limit of 128,000 tokens. It supports reasoning, tool use, and structured output. Accepted input modalities are Text, Image, and PDF. Providers report the capability flags inconsistently, so verify a specific feature against the host you plan to use. Its training knowledge cuts off at 2026-06.

Specs & pricing

Input / output per 1M tokens
Official price·Anthropic
$0.10 / $0.50
Blended $0.20 · Cache read $0.01
Lowest paid·NanoGPTGateway
$0.10 / $0.50
Blended $0.20 · 1.3× spread
Context
1,000,000
Output limit
128,000
Knowledge cutoff
2026-06
Released / updated
2026-10-07 / 2026-10-07
Capabilities
✓ Reasoning✓ Tool use✓ Structured outputTemperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImagePDF

Quality & performanceArtificial Analysis · Intelligence Index v4.3 · rep. tier Max

Intelligence43.4
Value217
Coding—
Agentic—
Output speed244 tok/s
TTFT341.02 s
Cost per task$0.21
Value formulaIQ 43.4 ÷ min blended $0.200 = 217
Reasoning tier → intelligence / speed (higher tier = stronger but slower)
LowIQ 29.4 · 185 tok/s
MediumIQ 34.5 · 155 tok/s
HighIQ 37.8 · 178 tok/s
XhighIQ 41.2 · 193 tok/s
MaxIQ 43.4 · 244 tok/s

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 21 providers21 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NanoGPT
anthropic/claude-haiku-5.5
Gateway$0.10$0.50$0.01$0.131,000,000128,000
AnthropicOfficialFirst-party$0.10
tiered >100K: $0.50
$0.50$0.01$0.131,000,000128,000
Kilo Gateway
anthropic/claude-haiku-5.5
Gateway$0.10$0.50$0.01$0.131,000,000128,000
OpenRouter
anthropic/claude-haiku-5.5
Gateway$0.10
tiered >100K: $0.50
$0.50$0.01$0.131,000,000128,000
AzureCloud$0.10
tiered >100K: $0.50
$0.50$0.01$0.131,000,000128,000
Azure Cognitive ServicesCloud$0.10
tiered >100K: $0.50
$0.50$0.01$0.131,000,000128,000
Vercel AI Gateway
anthropic/claude-haiku-5.5
Cloud$0.10
tiered >100K: $0.50
$0.50$0.01$0.131,000,000128,000
Vertex (Anthropic)
claude-haiku-5-5@default
Cloud$0.10
tiered >100K: $0.50
$0.50$0.01$0.131,000,000128,000
Merge GatewayGateway$0.10$0.50$0.01$0.131,000,000128,000
GitHub Copilot
claude-haiku-5.5
Cloud$0.10
tiered >100K: $0.50
$0.50$0.01$0.131,000,000128,000

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = loweffort = mediumeffort = higheffort = xhigheffort = maxToggle (on / off)effort = none

1 / 21 providers expose no reasoning control (reasoning_options: []).

Your usage cost

1NanoGPT$34.20
2Anthropic · Official$34.20
3Kilo Gateway$34.20
4OpenRouter$34.20
5Azure$34.20
6Azure Cognitive Services$34.20
Switch to NanoGPT to save $0/mo (0%). The gap is small, so staying on the official channel is fine.
Note: this is a gateway; verify availability and rate limits yourself.

Price historyone sample accumulated per data sync

Input list $0.10Output list $0.50Min blended $0.20

Price history accumulates from each data sync; currently only 1 sample(s) (2026-10-08). Each future sync adds a point, and once accumulated a line is drawn here.

Benchmark8 items

NameConditionsScoreMetricSource
GDPval-AAv2.11620EloSource ↗
AA-Briefcasev1.11578EloSource ↗
OSWorldvariant: offline subset · v2.172.4partial scoreSource ↗
Humanity's Last Examvariant: no tools45.9scoreSource ↗
Humanity's Last Examvariant: with tools57.4scoreSource ↗
Terminal-Benchv4.039.2scoreSource ↗
FrontierCodedataset: Main · v1.146.4scoreSource ↗
Chartographyvariant: no tools46.4scoreSource ↗

The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.

Artificial Analysis evaluations3 items

Humanity's Last Exam44.4%
SciCode55.0%
AA-LCR82.7%

Individual evaluations run by Artificial Analysis, on the same reasoning tier as the intelligence score above. Each benchmark has its own task set and harness, so rows are not comparable with one another. The Intelligence Index above draws on a different, newer set of evaluations.

Related models

Claude Haiku 5.5 (EU)same series$0.11 / $0.55Claude Haiku 5.5 (AU)same series$0.11 / $0.55Claude Haiku 5.5 (Global)same series$0.10 / $0.50Qwen3.7 Flashcheaper alternative$0.03 / $0.13Laguna S 2.1cheaper alternative$0 / $0Qwen Turbocheaper alternative$0.05 / $0.20Nemotron 3 Ultra (free)cheaper alternative$0 / $0
Data partly from models.dev (MIT) · Anthropic official docs ↗