DeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This is a rate-limited beta with limited capacity, intended for testing rather than production use. Assume prompts and responses are logged by the provider and may be used for model training or service improvement. Do not send sensitive or confidential data.
DeepSeek V4.1 Flash is currently listed from a single provider. Its reference price is $0.156 per 1M input tokens and $0.312 per 1M output tokens.
The context window is 1,000,000 tokens, with an output limit of 384,000 tokens. It supports reasoning, tool use, and structured output. Accepted input modalities are Text and Image.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NanoGPT | Gateway | $0.156 | $0.312 | $0.031 | — | 1,000,000 | 384,000 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Price history accumulates from each data sync; currently only 1 sample(s) (2026-09-09). Each future sync adds a point, and once accumulated a line is drawn here.