LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Qwen2.5 32B Instruct

alibaba·alibaba/qwen2-5-32b-instruct·GA·Open weights·qwen series
Qwen instruction model for multilingual chat, reasoning, and tool use

Qwen2.5 32B Instruct by alibaba is offered by 2 providers on this page. Its official list price is $0.70 per 1M input tokens and $2.80 per 1M output tokens. The lowest paid channel is Alibaba (China) at $0.287 / $0.861 per 1M, about 2.4× below the list price.

Artificial Analysis rates it 6.9 on the Intelligence Index. Against its lowest blended price of $0.43 per 1M, that is roughly 16 index points per dollar, which is the value ratio the leaderboards rank on.

The context window is 131,072 tokens, with an output limit of 8,192 tokens. It supports tool use. The weights are open, so it can also be self-hosted or served through a gateway of your choice. Its training knowledge cuts off at 2024-04.

Specs & pricing

Input / output per 1M tokens
Official price·Alibaba
$0.70 / $2.80
Blended $1.23 · Cache read —
Lowest paid·Alibaba (China)First-party
$0.29 / $0.86
Blended $0.43 · 2.4× spread
Context
131,072
Output limit
8,192
Knowledge cutoff
2024-04
Released / updated
2024-09 / 2024-09
Capabilities
Reasoning✓ Tool use? Structured output✓ TemperatureAttachments
Modalities
Text

Quality & performanceArtificial Analysis · Intelligence Index v4.3

Intelligence6.9
Value16
Coding—
Agentic—
Output speed—
TTFT—
Value formulaIQ 6.9 ÷ min blended $0.430 = 16

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 2 providers2 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
Alibaba (China)OfficialFirst-party$0.29$0.86——131,0728,192
AlibabaOfficialFirst-party$0.70$2.80——131,0728,192

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Alibaba (China) · Official$100.45
2Alibaba · Official$280.00
The cheapest paid channel is official.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$2.80$02026-08-052026-09-242026-08-05 · Input list $0.702026-08-06 · Input list $0.702026-08-07 · Input list $0.702026-08-08 · Input list $0.702026-08-09 · Input list $0.702026-08-10 · Input list $0.702026-08-11 · Input list $0.702026-08-12 · Input list $0.702026-08-13 · Input list $0.292026-09-24 · Input list $0.702026-08-05 · Output list $2.802026-08-06 · Output list $2.802026-08-07 · Output list $2.802026-08-08 · Output list $2.802026-08-09 · Output list $2.802026-08-10 · Output list $2.802026-08-11 · Output list $2.802026-08-12 · Output list $2.802026-08-13 · Output list $0.862026-09-24 · Output list $2.802026-08-05 · Min blended $0.432026-08-06 · Min blended $0.432026-08-07 · Min blended $0.432026-08-08 · Min blended $0.432026-08-09 · Min blended $0.432026-08-10 · Min blended $0.432026-08-11 · Min blended $0.432026-08-12 · Min blended $0.432026-08-13 · Min blended $0.432026-09-24 · Min blended $0.43

Artificial Analysis evaluations6 items

GPQA Diamond46.6%
Humanity's Last Exam4.0%
MMLU-Pro69.7%
LiveCodeBench24.8%
AIME11.0%
MATH-50080.5%

Individual evaluations run by Artificial Analysis, on the same reasoning tier as the intelligence score above. Each benchmark has its own task set and harness, so rows are not comparable with one another. The Intelligence Index above draws on a different, newer set of evaluations.

Related models

Qwen-Image-2.1same series$0 / $0Qwen 3.8 27B Hemingwaysame series$0.25 / $1.50Qwen 3.8 27B Cybersecuritysame series$0.10 / $0.60Qwen3 Coder Flashcheaper alternative$0.30 / $1.50Command Rcheaper alternative$0.15 / $0.60Gemma 3 12B ITcheaper alternative$0.27 / $0.27Qwen3 30B A3B Instruct 2507cheaper alternative$0.30 / $0.50
Data partly from models.dev (MIT) · Alibaba official docs ↗