← Model list
DeepSeek V4 Flash (Thinking)
deepseek·deepseek/deepseek-v4-flash-thinking·GA·Open weights·deepseek-flash series
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
At a glance
Reference price$0.14 / $0.28 per 1M
Cache read $0.0028 · NanoGPT
Lowest paid$0.14 / $0.28
NanoGPT Gateway
Context1,048,576
Output limit384,000
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
Modalities
Text
Knowledge cutoff2025-05
Released / updated2026-04-24 / 2026-04-24
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 1 providers1 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NanoGPT deepseek/deepseek-v4-flash:thinking | Gateway | $0.14 | $0.28 | $0.0028 | — | 1,048,576 | 384,000 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1NanoGPT$25.54
The cheapest paid channel is the only channel.
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
effort = noneeffort = higheffort = max
Related models
DeepSeek V4 Flash 0731 Fastsame series$0.35 / $0.70DeepSeek V4 Flash 0731 TEEsame series$0.14 / $0.28DeepSeek V4 Flash 0731same series$0.479 / $1.44Qwen3.7 Flashcheaper alternative$0.03 / $0.118Qwen Flashcheaper alternative$0.022 / $0.216Qwen Turbocheaper alternative$0.044 / $0.087Laguna S 2.1cheaper alternative$0 / $0
Price historyone sample accumulated per data sync
Input listOutput listMin blended