← Model list
DeepSeek V4 Flash
deepseek·deepseek/deepseek-v4-flash·GA·Open weights·deepseek-flash series
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
At a glance
Official price$0.14 / $0.28 per 1M
Cache read $0.0028 · DeepSeek
Lowest paid$0.063 / $0.125
UnoRouter Gateway · 7× spread
7 more $0 channels
Context1,000,000
⚠ Providers report 131,072–1,050,000; the table below is authoritative
Output limit384,000
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
Text
Knowledge cutoff2025-05
Released / updated2026-04-24 / 2026-04-24
Quality & performanceArtificial Analysis · Intelligence Index v4.1 · rep. tier Reasoning, Max Effort
Intelligence42.1
Value539
Coding56.2
Agentic33.7
Output speed—
TTFT—
Cost per task$0.0673
Value formulaIQ 42.1 ÷ min blended $0.078 = 539
Reasoning tier → intelligence / speed (higher tier = stronger but slower)
Non-reasoningIQ 29.3
Reasoning, High EffortIQ 39
Reasoning, Max EffortIQ 42.1
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 58 providers49 with public prices · 3 free · 4 subscription-covered
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Umans AI Coding Plan umans-deepseek-v4-flash-0731 | Gateway | Subscription | $0 | $0 | 1,048,576 ⚠ | 393,215 | ||
| Alibaba Token Plan | Gateway | Subscription | $0 | $0 | 1,000,000 | 384,000 | ||
| Kenari | Gateway | Free | — | — | 1,000,000 | 384,000 | ||
| UnoRouter deepseek-v4-flash:free | Gateway | Free | — | — | 1,000,000 | 384,000 | ||
| InferX | Gateway | Free | — | — | 1,000,000 | 100,000 | ||
| SCNet Token Plan DeepSeek-V4-Flash | Gateway | Subscription | $0 | — | 1,000,000 | 384,000 | ||
| Alibaba Token Plan (China) | Gateway | Subscription | $0 | $0 | 1,000,000 | 384,000 | ||
| UnoRouter | Gateway | $0.063 | $0.125 | — | — | 1,000,000 | 384,000 | |
| DigitalOcean deepseek-4-flash | Cloud | $0.068 | $0.168 | $0.017 | — | 1,048,576 ⚠ | 384,000 | |
| DevPass (LLM Gateway) | Gateway | $0.076 | $0.153 | $0.014 | — | 1,050,000 ⚠ | 384,000 | |
| DeepSeekOfficial | First-party | $0.14 | $0.28 | $0.0028 | — | 1,000,000 | 384,000 | |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1DevPass (LLM Gateway)$15.41
2DigitalOcean$15.85
3OpenRouter$16.85
4Deep Infra$18.36
5UnoRouter$18.75
6Pioneer$20.36
12DeepSeek · Official$25.54
7 more channels offer $0 (Umans AI Coding Plan, Alibaba Token Plan, Kenari etc.); free tiers usually have rate limits and no SLA, excluded from ranking.
Switch to DevPass (LLM Gateway) to save $10.13/mo (40%)
Note: this is a gateway; verify availability and rate limits yourself.
Benchmark1 items
| Name | Conditions | Score | Metric | Source |
|---|---|---|---|---|
| SWE-Bench Verified | — | 79 | resolved | Source ↗ |
The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.
Reasoning control
Toggle (on / off)effort = loweffort = higheffort = maxeffort = xhigheffort = noneeffort = minimaleffort = mediumbudget_tokens ≥ 1 ≤ 393,216
12 / 58 providers expose no reasoning control (reasoning_options: []).
Related models
DeepSeek V4 Flash 0731 Fastsame series$0.35 / $0.70DeepSeek V4 Flash 0731 TEEsame series$0.14 / $0.28DeepSeek V4 Flash 0731same series$0.479 / $1.44Qwen3.7 Flashcheaper alternative$0.03 / $0.118Qwen Flashcheaper alternative$0.022 / $0.216Qwen Turbocheaper alternative$0.044 / $0.087Laguna S 2.1cheaper alternative$0 / $0
Price historyone sample accumulated per data sync
Input listOutput listMin blended