LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

DeepSeek V4 Flash Flex

deepseek·deepseek/deepseek-v4-flash-flex·GA·Open weights·deepseek-flash series

Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work

At a glance
Reference price$0.091 / $0.182 per 1M
Cache read $0.018 · Neuralwatt
Lowest paid$0.091 / $0.182
Neuralwatt Gateway
Context1,048,560
Output limit65,536
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
Modalities
Text
Knowledge cutoff2025-05
Released / updated2026-04-24 / 2026-04-24

Quality & performance

Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NeuralwattGateway$0.091$0.182$0.018—1,048,56065,536

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Neuralwatt$18.56
The cheapest paid channel is the only channel.

Benchmark

No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.

Reasoning control

effort = noneeffort = higheffort = max

Related models

DeepSeek V4 Flash Vision Expsame series$0.14 / $0.28DeepSeek V4 Flash 0731 Fastsame series$0.35 / $0.70DeepSeek V4 Flash 0731 TEEsame series$0.44 / $1.32Qwen3.7 Flashcheaper alternative$0.03 / $0.118Qwen Turbocheaper alternative$0.044 / $0.087Laguna S 2.1cheaper alternative$0 / $0

Price historyone sample accumulated per data sync

Input list $0.091Output list $0.182Min blended $0.114

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-27). Each future sync adds a point, and once accumulated a line is drawn here.

Data from models.dev (MIT) · Neuralwatt official docs ↗