Fast Grok model for responsive chat, reasoning, and tool-assisted work
grok-4-fast-non-reasoning by xai is offered by 6 providers on this page. Public prices are shown for 5 of them. Its reference price is $0.20 per 1M input tokens and $0.50 per 1M output tokens.
The context window is 2,000,000 tokens, with an output limit of 2,000,000 tokens. It supports tool use and structured output. Accepted input modalities are Text, Image, Audio, and Video. Providers report the capability flags inconsistently, so verify a specific feature against the host you plan to use. Its training knowledge cuts off at 2025-06.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Jiekou.AI | Gateway | $0.18 | $0.45 | — | — | 2,000,000 | 2,000,000 | |
| Poe | Gateway | $0.20 | $0.50 | $0.05 | — | 2,000,000 | 128,000 | |
| Abacus | Gateway | $0.20 | $0.50 | — | — | 2,000,000 | 16,384 | |
| 302.AI | Gateway | $0.20 | $0.50 | — | — | 2,000,000 | 30,000 | |
| Helicone | Gateway | $0.20 | $0.50 | $0.05 | — | 2,000,000 | 2,000,000 | |
| Qiniu | Gateway | — | — | — | — | 2,000,000 | 2,000,000 | hostTable.undisclosed |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.