LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

DeepSeek V4.1 Flash Thinking

deepseek·deepseek/deepseek-v4.1-flash-thinking·GA·Open weights·deepseek-flash series·NEW
DeepSeek V4.1 Flash model for reasoning and agentic coding

Specs & pricing

Input / output per 1M tokens
Reference price·NanoGPT
$0.30 / $1.20
Blended $0.525 · Cache read $0.006
Lowest paid·NanoGPTGateway
$0.30 / $1.20
Blended $0.525
Context
1,000,000
Output limit
384,000
Knowledge cutoff
2025-05
Released / updated
2026-09-10 / 2026-09-10
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImage

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NanoGPT
deepseek/deepseek-v4.1-flash:thinking
Gateway$0.30$1.20$0.006—1,000,000384,000

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1NanoGPT$84.72
The cheapest paid channel is the only channel.

Reasoning control

effort = noneeffort = loweffort = higheffort = max

Related models

DeepSeek V4.1 Flashsame series$0.15 / $0.60DeepSeek V4 Flash Vision Expsame series$0.15 / $0.60DeepSeek V4 Flash 0731 Fastsame series$0.35 / $0.70DeepSeek V4 Flashcheaper alternative$0.15 / $0.60GLM-5.3-Flashcheaper alternative$0.075 / $0.25MiMo-V2.5cheaper alternative$0.14 / $0.28Gemini 2.5 Flash-Litecheaper alternative$0.10 / $0.40

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$1.20$02026-09-102026-09-112026-09-10 · Input list $0.1562026-09-11 · Input list $0.302026-09-10 · Output list $0.3122026-09-11 · Output list $1.202026-09-10 · Min blended $0.1952026-09-11 · Min blended $0.525
Data partly from models.dev (MIT) · NanoGPT official docs ↗