← Model list

DeepSeek V4 Flash (Alibaba Cloud)

deepseek·deepseek/deepseek-v4-flash-10·GA·Open weights·deepseek-flash series

Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work

At a glance
Reference price$0.20 / $0.40 per 1M
Cache read $0.04 · LLM Gateway
Lowest paid$0.14 / $0.28
AIHubMix Gateway · 1.4× spread
Context1,000,000
Output limit393,216
Capabilities
ReasoningTool useStructured outputTemperatureAttachments
Modalities
Text
Knowledge cutoff2025-05
Released / updated2026-04-24 / 2026-04-24

Quality & performance

Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.

Available at 2 providers2 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
AIHubMix
alicloud-deepseek-v4-flash
Gateway$0.14$0.28$0.0281,000,000384,000
LLM Gateway
alibaba/deepseek-v4-flash
Gateway$0.20$0.40$0.041,000,000393,216

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1AIHubMix$28.56
2LLM Gateway$40.80
The cheapest paid channel is the only channel.

Benchmark

No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.

Reasoning control

Toggle (on / off)effort = noneeffort = minimaleffort = loweffort = mediumeffort = higheffort = xhigheffort = max

Related models

Price historyone sample accumulated per data sync

Input list $0.20Output list $0.40Min blended $0.175

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-21). Each future sync adds a point, and once accumulated a line is drawn here.