LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

DeepSeek V4 Flash

deepseek·deepseek/deepseek-v4-flash·Deprecated·Open weights·deepseek-flash series
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work

DeepSeek V4 Flash is offered by 75 providers on this page. Public prices are shown for 63 of them. Its official list price is $0.15 per 1M input tokens and $0.60 per 1M output tokens. The lowest paid channel is Merge Gateway at $0.035 / $0.07 per 1M, about 15× below the list price. That channel is a third-party gateway, so confirm its availability and rate limits before depending on it. It also has 5 free ($0) channels and 5 covered by a paid subscription; free tiers usually carry rate limits, and subscription-covered access bills $0 per token only after the subscription fee.

The context window is 1,000,000 tokens at the reference host, but hosts report different limits, from 163,840 to 1,050,000, so the usable window depends on the provider you pick. It supports reasoning, tool use, and structured output. Accepted input modalities are Text and Image. The weights are open, so it can also be self-hosted or served through a gateway of your choice. Providers report the capability flags inconsistently, so verify a specific feature against the host you plan to use. Its training knowledge cuts off at 2025-05.

Specs & pricing

Input / output per 1M tokens
Official price·DeepSeek
$0.15 / $0.60
Blended $0.26 · Cache read $0.003
Lowest paid·Merge GatewayGateway
$0.035 / $0.07
Blended $0.044 · 14.7× spread
5 more $0 channels
Context
1,000,000
Output limit
393,216
Knowledge cutoff
2025-05
Released / updated
2026-04-24 / 2026-04-24
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
TextImage
Weights
Hugging Face

Available at 75 providers63 with public prices · 5 free · 5 subscription-covered

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
SenseNova (China)GatewayFree$0—1,000,00065,536
Alibaba Token PlanGatewaySubscription$0$01,000,000384,000
Umans AI Coding Plan
umans-deepseek-v4-flash-0731
GatewaySubscription$0$01,048,576 ⚠393,215
Alibaba Token Plan (China)GatewaySubscription$0$01,000,000384,000
PendraGatewayFree——1,000,000384,000
UnoRouter
deepseek-v4-flash:free
GatewayFree——1,000,000384,000
Volcengine Ark Coding PlanGatewaySubscription$0—1,000,000384,000
SCNet Token Plan
DeepSeek-V4-Flash
GatewaySubscription$0—1,000,000384,000
InferXGatewayFree——1,000,000100,000
KenariGatewayFree——1,000,000384,000
DeepSeekOfficialFirst-party$0.15$0.60$0.003—1,000,000393,216deprecated

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = noneeffort = highToggle (on / off)effort = maxeffort = loweffort = xhigheffort = minimaleffort = mediumbudget_tokens ≥ 128 ≤ 32,768

Interleaved thinking (reasoning between tool calls) is declared by 49 of 75 providers.

11 / 75 providers expose no reasoning control (reasoning_options: []).

Your usage cost

1Merge Gateway$7.14
2DevPass (LLM Gateway)$12.44
3LLM Gateway$12.44
4LLM Gateway$15.41
5LLM Gateway$17.32
6Deep Infra$18.36
48DeepSeek · Official$42.36

10 more channels offer $0 (SenseNova (China), Alibaba Token Plan, Umans AI Coding Plan etc.); free tiers usually have rate limits and no SLA, excluded from ranking.

Switch to Merge Gateway to save $35.22/mo (83%)
Note: this is a gateway; verify availability and rate limits yourself.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$0.60$02026-08-052026-10-042026-08-05 · Input list $0.142026-08-06 · Input list $0.142026-08-07 · Input list $0.142026-08-08 · Input list $0.142026-08-09 · Input list $0.142026-08-10 · Input list $0.142026-08-11 · Input list $0.142026-08-12 · Input list $0.142026-08-13 · Input list $0.142026-08-19 · Input list $0.142026-08-20 · Input list $0.142026-08-23 · Input list $0.142026-08-24 · Input list $0.142026-08-25 · Input list $0.142026-08-27 · Input list $0.142026-09-06 · Input list $0.142026-09-11 · Input list $0.152026-10-03 · Input list $0.152026-10-04 · Input list $0.152026-08-05 · Output list $0.282026-08-06 · Output list $0.282026-08-07 · Output list $0.282026-08-08 · Output list $0.282026-08-09 · Output list $0.282026-08-10 · Output list $0.282026-08-11 · Output list $0.282026-08-12 · Output list $0.282026-08-13 · Output list $0.282026-08-19 · Output list $0.282026-08-20 · Output list $0.282026-08-23 · Output list $0.282026-08-24 · Output list $0.282026-08-25 · Output list $0.282026-08-27 · Output list $0.282026-09-06 · Output list $0.282026-09-11 · Output list $0.602026-10-03 · Output list $0.602026-10-04 · Output list $0.602026-08-05 · Min blended $0.0782026-08-06 · Min blended $0.0782026-08-07 · Min blended $0.0782026-08-08 · Min blended $0.0782026-08-09 · Min blended $0.0782026-08-10 · Min blended $0.0782026-08-11 · Min blended $0.0782026-08-12 · Min blended $0.0782026-08-13 · Min blended $0.0782026-08-19 · Min blended $0.062026-08-20 · Min blended $0.0782026-08-23 · Min blended $0.0682026-08-24 · Min blended $0.0732026-08-25 · Min blended $0.0782026-08-27 · Min blended $0.0642026-09-06 · Min blended $0.0632026-09-11 · Min blended $0.0442026-10-03 · Min blended $0.0352026-10-04 · Min blended $0.044

Benchmark22 items

NameConditionsScoreMetricSource
SWE-Bench Verified—79resolvedSource ↗
MMLU-Provariant: preview checkpoint; max effort86.2EMSource ↗
SimpleQA-Verifiedvariant: preview checkpoint; max effort34.1pass@1Source ↗
Chinese SimpleQAvariant: preview checkpoint; max effort78.9pass@1Source ↗
GPQA Diamondvariant: preview checkpoint; max effort88.1pass@1Source ↗
Humanity's Last Examvariant: preview checkpoint; max effort; without tools34.8pass@1Source ↗
LiveCodeBenchvariant: preview checkpoint; max effort91.6pass@1Source ↗
Codeforcesvariant: preview checkpoint; max effort3052ratingSource ↗
HMMTvariant: preview checkpoint; max effort · dataset: February 202694.8pass@1Source ↗
IMOAnswerBenchvariant: preview checkpoint; max effort88.4pass@1Source ↗
MathArena Apexvariant: preview checkpoint; max effort33pass@1Source ↗
MathArena Apex Shortlistvariant: preview checkpoint; max effort85.7pass@1Source ↗
MRCRvariant: preview checkpoint; max effort · dataset: 1M context78.7MMRSource ↗
CorpusQAvariant: preview checkpoint; max effort · dataset: 1M context60.5accuracySource ↗
Terminal-Benchvariant: preview checkpoint; max effort · v2.056.9accuracySource ↗
SWE-Bench Provariant: preview checkpoint; max effort52.6resolvedSource ↗
SWE-Bench Multilingualvariant: preview checkpoint; max effort73.3resolvedSource ↗
BrowseCompvariant: preview checkpoint; max effort73.2pass@1Source ↗
Humanity's Last Examvariant: preview checkpoint; max effort; with tools45.1pass@1Source ↗
MCP Atlasvariant: preview checkpoint; max effort · dataset: public69pass@1Source ↗
GDPval-AAvariant: preview checkpoint; max effort1395EloSource ↗
Toolathlonvariant: preview checkpoint; max effort47.8pass@1Source ↗

The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.

Related models

DeepSeek Flash Latestsame series$0.003 / $2.40DeepSeek V4.1 Flashsame series$0.15 / $0.60deepseek-ai/DeepSeek-V4.1-Flash-Fastsame series$0.60 / $2.40Qwen3.7 Flashcheaper alternative$0.03 / $0.13Muse Spark 1.3 Contributorcheaper alternative$0.10 / $0.20Muse Spark 1.2 Contributorcheaper alternative$0.10 / $0.20Qwen Flashcheaper alternative$0.05 / $0.40
Data partly from models.dev (MIT) · DeepSeek official docs ↗