LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

DeepSeek V4.1 Flash Flex

deepseek·deepseek/deepseek-v4.1-flash-flex·GA·Open weights·deepseek-flash series·NEW
DeepSeek V4.1 Flash model for reasoning and agentic coding

Specs & pricing

Input / output per 1M tokens
Reference price·Neuralwatt
$0.098 / $0.39
Blended $0.17 · Cache read $0.0098
Lowest paid·NeuralwattGateway
$0.098 / $0.39
Blended $0.17
Context
1,048,560
Output limit
393,216
Knowledge cutoff
2025-05
Released / updated
2026-09-10 / 2026-09-10
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImage

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NeuralwattGateway$0.098$0.39$0.0098—1,048,560393,216

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = noneeffort = loweffort = higheffort = xhigheffort = max

Interleaved thinking (reasoning between tool calls) is declared by 1 of 1 providers.

Your usage cost

1Neuralwatt$28.47
The cheapest paid channel is the only channel.

Price historyone sample accumulated per data sync

Input list $0.098Output list $0.39Min blended $0.17

Price history accumulates from each data sync; currently only 1 sample(s) (2026-09-22). Each future sync adds a point, and once accumulated a line is drawn here.

Related models

DeepSeek Flash Latestsame series$0.003 / $2.40DeepSeek V4.1 Flashsame series$0.15 / $0.60deepseek-ai/DeepSeek-V4.1-Flash-Fastsame series$0.60 / $2.40Qwen3.7 Flashcheaper alternative$0.03 / $0.13Laguna S 2.1cheaper alternative$0 / $0Qwen Turbocheaper alternative$0.05 / $0.20Nemotron 3 Ultra (free)cheaper alternative$0 / $0
Data partly from models.dev (MIT) · Neuralwatt official docs ↗