LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

DeepSeek V4 Pro

deepseek·deepseek/deepseek-v4-pro·Deprecated·Open weights·deepseek-thinking series
Open MoE flagship with million-token context for coding and long agent runs

DeepSeek V4 Pro is offered by 78 providers on this page. Public prices are shown for 69 of them. Its official list price is $0.66 per 1M input tokens and $1.98 per 1M output tokens. The lowest paid channel is OpenRouter at $0.209 / $0.418 per 1M, about 12× below the list price. That channel is a third-party gateway, so confirm its availability and rate limits before depending on it. It also has 3 free ($0) channels and 4 covered by a paid subscription; free tiers usually carry rate limits, and subscription-covered access bills $0 per token only after the subscription fee.

The context window is 1,000,000 tokens at the reference host, but hosts report different limits, from 128,000 to 1,050,000, so the usable window depends on the provider you pick. It supports reasoning, tool use, and structured output. The weights are open, so it can also be self-hosted or served through a gateway of your choice. Providers report the capability flags inconsistently, so verify a specific feature against the host you plan to use. Its training knowledge cuts off at 2025-05.

Specs & pricing

Input / output per 1M tokens
Official price·DeepSeek
$0.66 / $1.98
Blended $0.99 · Cache read $0.022
Lowest paid·OpenRouterGateway
$0.21 / $0.42
Blended $0.26 · 11.5× spread
3 more $0 channels
Context
1,000,000
Output limit
393,216
Knowledge cutoff
2025-05
Released / updated
2026-04-24 / 2026-04-24
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
Text
Weights
Hugging Face

Available at 78 providers69 with public prices · 3 free · 4 subscription-covered

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
SenseNova (China)GatewayFree$0—1,048,576 ⚠65,536
Alibaba Token PlanGatewaySubscription$0$01,000,000384,000
Alibaba Token Plan (China)GatewaySubscription$0$01,000,000384,000
UnoRouter
deepseek-v4-pro:free
GatewayFree——1,000,000384,000
Volcengine Ark Coding PlanGatewaySubscription$0—1,000,000384,000
SCNet Token Plan
DeepSeek-V4-Pro
GatewaySubscription$0—1,000,000384,000
KenariGatewayFree——1,000,000384,000
OpenRouterGateway$0.21$0.42$0.017—1,048,576 ⚠384,000
routing.runGateway$0.35$0.70——1,000,00064,000
CrofAIGateway$0.35$0.80$0.003—1,000,000131,072
DeepSeekOfficialFirst-party$0.66$1.98$0.022—1,000,000393,216

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

Toggle (on / off)effort = higheffort = maxeffort = xhigheffort = noneeffort = loweffort = mediumeffort = minimalbudget_tokens ≥ 1 ≤ 393,216

Interleaved thinking (reasoning between tool calls) is declared by 52 of 78 providers.

11 / 78 providers expose no reasoning control (reasoning_options: []).

Your usage cost

1OpenRouter$39.67
2CrofAI$68.36
3Auriko$78.74
4ZenMux$78.74
5Hugging Face$78.74
6Nvidia$78.74
19DeepSeek · Official$154.44

7 more channels offer $0 (SenseNova (China), Alibaba Token Plan, Alibaba Token Plan (China) etc.); free tiers usually have rate limits and no SLA, excluded from ranking.

Switch to OpenRouter to save $114.77/mo (74%)
Note: this is a gateway; verify availability and rate limits yourself.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$1.98$02026-08-052026-10-032026-08-05 · Input list $0.442026-08-06 · Input list $0.442026-08-07 · Input list $0.442026-08-08 · Input list $0.442026-08-09 · Input list $0.442026-08-10 · Input list $0.442026-08-11 · Input list $0.442026-08-12 · Input list $0.442026-08-13 · Input list $0.442026-10-01 · Input list $0.442026-10-02 · Input list $0.442026-10-03 · Input list $0.662026-08-05 · Output list $0.872026-08-06 · Output list $0.872026-08-07 · Output list $0.872026-08-08 · Output list $0.872026-08-09 · Output list $0.872026-08-10 · Output list $0.872026-08-11 · Output list $0.872026-08-12 · Output list $0.872026-08-13 · Output list $0.872026-10-01 · Output list $0.872026-10-02 · Output list $0.872026-10-03 · Output list $1.982026-08-05 · Min blended $0.442026-08-06 · Min blended $0.442026-08-07 · Min blended $0.442026-08-08 · Min blended $0.442026-08-09 · Min blended $0.442026-08-10 · Min blended $0.442026-08-11 · Min blended $0.442026-08-12 · Min blended $0.442026-08-13 · Min blended $0.442026-10-01 · Min blended $0.272026-10-02 · Min blended $0.262026-10-03 · Min blended $0.26

Benchmark26 items

NameConditionsScoreMetricSource
SWE-Bench Verified—80.6resolvedSource ↗
Artificial Analysis Coding Agent Indexharness: Claude Code · variant: high50.1average pass@1Source ↗
SWE-Atlas Codebase QnAharness: Claude Code · variant: high67.8pass@1Source ↗
SWE-Bench Proharness: Claude Code · variant: high · dataset: hard-aa18pass@1Source ↗
Terminal-Benchharness: Claude Code · variant: high · v2.164.7pass@1Source ↗
MMLU-Provariant: preview checkpoint; max effort87.5EMSource ↗
SimpleQA-Verifiedvariant: preview checkpoint; max effort57.9pass@1Source ↗
Chinese SimpleQAvariant: preview checkpoint; max effort84.4pass@1Source ↗
GPQA Diamondvariant: preview checkpoint; max effort90.1pass@1Source ↗
Humanity's Last Examvariant: preview checkpoint; max effort; without tools37.7pass@1Source ↗
LiveCodeBenchvariant: preview checkpoint; max effort93.5pass@1Source ↗
Codeforcesvariant: preview checkpoint; max effort3206ratingSource ↗
HMMTvariant: preview checkpoint; max effort · dataset: February 202695.2pass@1Source ↗
IMOAnswerBenchvariant: preview checkpoint; max effort89.8pass@1Source ↗
MathArena Apexvariant: preview checkpoint; max effort38.3pass@1Source ↗
MathArena Apex Shortlistvariant: preview checkpoint; max effort90.2pass@1Source ↗
MRCRvariant: preview checkpoint; max effort · dataset: 1M context83.5MMRSource ↗
CorpusQAvariant: preview checkpoint; max effort · dataset: 1M context62accuracySource ↗
Terminal-Benchvariant: preview checkpoint; max effort · v2.067.9accuracySource ↗
SWE-Bench Provariant: preview checkpoint; max effort55.4resolvedSource ↗
SWE-Bench Multilingualvariant: preview checkpoint; max effort76.2resolvedSource ↗
BrowseCompvariant: preview checkpoint; max effort83.4pass@1Source ↗
Humanity's Last Examvariant: preview checkpoint; max effort; with tools48.2pass@1Source ↗
MCP Atlasvariant: preview checkpoint; max effort · dataset: public73.6pass@1Source ↗
GDPval-AAvariant: preview checkpoint; max effort1554EloSource ↗
Toolathlonvariant: preview checkpoint; max effort51.8pass@1Source ↗

The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.

Related models

DeepSeek V4 Pro 0813 (Databricks)same series$1.32 / $3.96DeepSeek V4 Pro 0813 (EU)same series$1.75 / $3.50DeepSeek V4 Pro 0813 Thinkingsame series$1.10 / $2.50GLM-5.3-Flashcheaper alternative$0.15 / $0.50DeepSeek V4.1 Flashcheaper alternative$0.15 / $0.60GPT-5.6 Lunacheaper alternative$0.20 / $1.20Qwen3.8 Flashcheaper alternative$0.15 / $0.47
Data partly from models.dev (MIT) · DeepSeek official docs ↗