LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

DeepSeek R1 Distill Qwen 14B

alibaba·alibaba/deepseek-r1-distill-qwen-14b·GA·Closed·qwen series
Qwen instruction model for multilingual chat, reasoning, and tool use

DeepSeek R1 Distill Qwen 14B by alibaba is currently listed from a single provider. Its official list price is $0.144 per 1M input tokens and $0.431 per 1M output tokens.

Artificial Analysis rates it 7.8 on the Intelligence Index. Against its lowest blended price of $0.216 per 1M, that is roughly 36 index points per dollar, which is the value ratio the leaderboards rank on.

The context window is 32,768 tokens, with an output limit of 16,384 tokens. It supports reasoning and tool use.

Specs & pricing

Input / output per 1M tokens
Official price·Alibaba (China)
$0.14 / $0.43
Blended $0.22 · Cache read —
Lowest paid·Alibaba (China)First-party
$0.14 / $0.43
Blended $0.22
Context
32,768
Output limit
16,384
Knowledge cutoff
—
Released / updated
2025-01-01 / 2025-01-01
Capabilities
✓ Reasoning✓ Tool use? Structured output✓ TemperatureAttachments
Modalities
Text

Quality & performanceArtificial Analysis · Intelligence Index v4.3

Intelligence7.8
Value36
Coding—
Agentic—
Output speed—
TTFT—
Value formulaIQ 7.8 ÷ min blended $0.216 = 36

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
Alibaba (China)OfficialFirst-party$0.14$0.43——32,76816,384

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

Upstream provides no control info

1 / 1 providers expose no reasoning control (reasoning_options: []).

Your usage cost

1Alibaba (China) · Official$50.35
The cheapest paid channel is official.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$0.43$02026-08-052026-10-102026-08-05 · Input list $0.142026-08-06 · Input list $0.142026-08-07 · Input list $0.142026-08-08 · Input list $0.142026-08-09 · Input list $0.142026-08-10 · Input list $0.142026-08-11 · Input list $0.142026-08-12 · Input list $0.142026-08-13 · Input list $0.142026-10-10 · Input list $0.142026-08-05 · Output list $0.432026-08-06 · Output list $0.432026-08-07 · Output list $0.432026-08-08 · Output list $0.432026-08-09 · Output list $0.432026-08-10 · Output list $0.432026-08-11 · Output list $0.432026-08-12 · Output list $0.432026-08-13 · Output list $0.432026-10-10 · Output list $0.432026-08-05 · Min blended $0.152026-08-06 · Min blended $0.152026-08-07 · Min blended $0.152026-08-08 · Min blended $0.152026-08-09 · Min blended $0.152026-08-10 · Min blended $0.152026-08-11 · Min blended $0.152026-08-12 · Min blended $0.152026-08-13 · Min blended $0.152026-10-10 · Min blended $0.22

Artificial Analysis evaluations9 items

GPQA Diamond48.4%
Humanity's Last Exam4.1%
MMLU-Pro74.0%
LiveCodeBench37.6%
AIME 202555.7%
AIME66.7%
MATH-50094.9%
AA-LCR10.3%
IFBench22.1%

Individual evaluations run by Artificial Analysis, on the same reasoning tier as the intelligence score above. Each benchmark has its own task set and harness, so rows are not comparable with one another. The Intelligence Index above draws on a different, newer set of evaluations.

Related models

Qwen3 30B A3B Instruct 2507same series$0.30 / $0.50Qwen 3.8 27B Antislopsame series$0.10 / $0.60Qwen 3.8 27B DarkIdolsame series$0.10 / $0.60Qwen3.7 Flashcheaper alternative$0.03 / $0.13Nemotron 3.5 Lightning 30B A3Bcheaper alternative$0 / $0Muse Spark 1.3 Contributorcheaper alternative$0.10 / $0.20Muse Spark 1.2 Contributorcheaper alternative$0.10 / $0.20
Data partly from models.dev (MIT) · Alibaba (China) official docs ↗