← Model list

DeepSeek V4 Flash 0731 Fast

deepseek·deepseek/deepseek-v4-flash-0731-fast·GA·Open weights·deepseek-flash series·NEW

Fast DeepSeek model for efficient chat, coding help, and agent loops

At a glance
Reference price$0.35 / $0.70 per 1M
Cache read $0.087 · Venice AI
Lowest paid$0.35 / $0.70
Venice AI Gateway
Context1,000,000
Output limit32,768
Capabilities
ReasoningTool useStructured outputTemperatureAttachments
Modalities
Text
Knowledge cutoff2025-05
Released / updated2026-08-09 / 2026-08-11

Quality & performance

Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
Venice AIGateway$0.35$0.70$0.0871,000,00032,768

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Venice AI$73.50
The cheapest paid channel is the only channel.

Benchmark

No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.

Reasoning control

effort = noneeffort = loweffort = higheffort = max

Related models

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$0.70$02026-08-122026-08-132026-08-12 · Input list $0.352026-08-13 · Input list $0.352026-08-12 · Output list $0.702026-08-13 · Output list $0.702026-08-12 · Min blended $0.4382026-08-13 · Min blended $0.438