← Model list DeepSeek V4 Flash 0731 Fast deepseek ·deepseek/deepseek-v4-flash-0731-fast ·GA ·Open weights ·deepseek-flash series ·NEW
WatchCopy ID + Add to compare Fast DeepSeek model for efficient chat, coding help, and agent loops
Specs & pricing Input / output per 1M tokens $0.28 / $1.40
Blended $0.56 · Cache read $0.07
Lowest paid · Venice AI Gateway $0.35 / $0.70
Blended $0.44 · 1.2× spread
Released / updated
2026-08-09 / 2026-08-11
Capabilities
✓ Reasoning ✓ Tool use ✓ Structured output ✓ Temperature Attachments
Available at 2 providers2 with public prices Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Reasoning control effort = none effort = low effort = high effort = max Toggle (on / off)
Interleaved thinking (reasoning between tool calls) is declared by 1 of 2 providers.
Your usage cost 1 Venice AI $73.50
2 AIHubMix $100.80
The cheapest paid channel is the only channel.
Price historyone sample accumulated per data sync Input list Output list Min blended
$1.40 $0 2026-08-12 2026-09-22 2026-08-12 · Input list $0.35 2026-08-13 · Input list $0.35 2026-09-22 · Input list $0.28 2026-08-12 · Output list $0.70 2026-08-13 · Output list $0.70 2026-09-22 · Output list $1.40 2026-08-12 · Min blended $0.44 2026-08-13 · Min blended $0.44 2026-09-22 · Min blended $0.44 Related models