LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

DeepSeek R1 0528 Qwen3 8B

deepseek·deepseek/deepseek-r1-0528-qwen3-8b·GA·Open weights
DeepSeek reasoning model for multi-step analysis, math, coding, and tools

DeepSeek R1 0528 Qwen3 8B is currently listed from a single provider. Its reference price is $0.06 per 1M input tokens and $0.09 per 1M output tokens.

Artificial Analysis rates it 8.1 on the Intelligence Index. Against its lowest blended price of $0.068 per 1M, that is roughly 120 index points per dollar, which is the value ratio the leaderboards rank on.

The context window is 128,000 tokens, with an output limit of 32,000 tokens. It supports reasoning. The weights are open, so it can also be self-hosted or served through a gateway of your choice.

Specs & pricing

Input / output per 1M tokens
Reference price·NovitaAI
$0.06 / $0.09
Blended $0.068 · Cache read —
Lowest paid·NovitaAIGateway
$0.06 / $0.09
Blended $0.068
Context
128,000
Output limit
32,000
Knowledge cutoff
—
Released / updated
2025-05-29 / 2025-05-29
Capabilities
✓ ReasoningTool use? Structured output✓ TemperatureAttachments
Modalities
Text

Quality & performanceArtificial Analysis · Intelligence Index v4.3

Intelligence8.1
Value120
Coding—
Agentic—
Output speed—
TTFT—
Value formulaIQ 8.1 ÷ min blended $0.068 = 120

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NovitaAIGateway$0.06$0.09——128,00032,000

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

Upstream provides no control info

1 / 1 providers expose no reasoning control (reasoning_options: []).

Your usage cost

1NovitaAI$16.50
The cheapest paid channel is the only channel.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$0.09$02026-08-052026-08-132026-08-05 · Input list $0.062026-08-06 · Input list $0.062026-08-07 · Input list $0.062026-08-08 · Input list $0.062026-08-09 · Input list $0.062026-08-10 · Input list $0.062026-08-11 · Input list $0.062026-08-12 · Input list $0.062026-08-13 · Input list $0.062026-08-05 · Output list $0.092026-08-06 · Output list $0.092026-08-07 · Output list $0.092026-08-08 · Output list $0.092026-08-09 · Output list $0.092026-08-10 · Output list $0.092026-08-11 · Output list $0.092026-08-12 · Output list $0.092026-08-13 · Output list $0.092026-08-05 · Min blended $0.0682026-08-06 · Min blended $0.0682026-08-07 · Min blended $0.0682026-08-08 · Min blended $0.0682026-08-09 · Min blended $0.0682026-08-10 · Min blended $0.0682026-08-11 · Min blended $0.0682026-08-12 · Min blended $0.0682026-08-13 · Min blended $0.068

Artificial Analysis evaluations11 items

GPQA Diamond61.2%
Humanity's Last Exam5.9%
MMLU-Pro73.9%
LiveCodeBench51.3%
Terminal-Bench Hard1.5%
AIME 202563.7%
AIME65.0%
MATH-50093.2%
τ²-Bench Telecom0.0%
AA-LCR15.0%
IFBench19.9%

Individual evaluations run by Artificial Analysis, on the same reasoning tier as the intelligence score above. Each benchmark has its own task set and harness, so rows are not comparable with one another. The Intelligence Index above draws on a different, newer set of evaluations.

Related models

Hy3cheaper alternative$0 / $0Nemotron 3.5 Lightning 30B A3Bcheaper alternative$0 / $0Laguna S 2.1cheaper alternative$0 / $0GLM-4.6V-Flashcheaper alternative$0 / $0
Data partly from models.dev (MIT) · NovitaAI official docs ↗