← Model list
Qwen3 8B
alibaba·alibaba/qwen3-8b·GA·Open weights·qwen series
Qwen instruction model for multilingual chat, reasoning, and tool use
At a glance
Official price$0.072 / $0.287 per 1M
Cache read — · Alibaba (China)
Lowest paid$0.035 / $0.138
NovitaAI Gateway · 13.4× spread
Context131,072
⚠ Providers report 40,960–131,072; the table below is authoritative
Output limit8,192
Capabilities
✓ Reasoning✓ Tool useStructured output✓ TemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
Text
Knowledge cutoff2025-03-31
Released / updated2025-04 / 2025-04-29
Quality & performanceArtificial Analysis · Intelligence Index v4.1 · rep. tier Reasoning
Intelligence8.3
Value137
Coding9
Agentic1.6
Output speed39 tok/s
TTFT3.76 s
Value formulaIQ 8.3 ÷ min blended $0.061 = 137
Reasoning tier → intelligence / speed (higher tier = stronger but slower)
Non-reasoningIQ 4.8 · 39 tok/s
ReasoningIQ 8.3 · 39 tok/s
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 6 providers6 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NovitaAI qwen/qwen3-8b-fp8 | Gateway | $0.035 | $0.138 | — | — | 128,000 ⚠ | 20,000 | |
| Alibaba (China)Official | First-party | $0.072 | $0.287 | — | — | 131,072 | 8,192 | |
| Pioneer Qwen/Qwen3-8B | Gateway | $0.20 | $0.20 | $0.20 | $0.20 | 40,960 ⚠ | 40,960 | |
| OpenRouter | Gateway | $0.117 | $0.455 | — | — | 131,072 | 8,192 | |
| AlibabaOfficial | First-party | $0.18 | $0.70 | — | — | 131,072 | 8,192 | |
| NanoGPT qwen/Qwen3-8B | Gateway | $0.47 | $0.47 | $0.235 | — | 41,000 ⚠ | 32,768 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1NovitaAI$13.90
2Alibaba (China) · Official$28.75
3OpenRouter$46.15
4Pioneer$50.00
5Alibaba · Official$71.00
6NanoGPT$89.30
Switch to NovitaAI to save $14.85/mo (52%)
Note: this is a gateway; verify availability and rate limits yourself.
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
Toggle (on / off)budget_tokenseffort = loweffort = mediumeffort = high
1 / 6 providers expose no reasoning control (reasoning_options: []).
Related models
Qwen3.5 0.8Bsame series$0.06 / $0.12Qwen3.5 4Bsame series$0.10 / $0.20Qwen3.8 27B TEEsame series$0.40 / $3.00GLM-4.7-Flashcheaper alternative$0 / $0Hy3cheaper alternative$0 / $0Nemotron 3 Nano 30B A3Bcheaper alternative$0 / $0Qwen3.7 Flashcheaper alternative$0.03 / $0.118
Price historyone sample accumulated per data sync
Input listOutput listMin blended