LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

X-Ai/Grok 4.1 Fast Non Reasoning

xai·xai/grok-4.1-fast-non-reasoning·GA·Closed·grok series
Fast Grok model for responsive chat, reasoning, and tool-assisted work

Specs & pricing

Input / output per 1M tokens
Reference price·Helicone
$0.20 / $0.50
Blended $0.28 · Cache read $0.05
Lowest paid·HeliconeGateway
$0.20 / $0.50
Blended $0.28
Context
2,000,000
Output limit
30,000
Knowledge cutoff
2025-11
Released / updated
2025-12-19 / 2025-12-19
Capabilities
Reasoning✓ Tool useStructured output✓ Temperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImageAudioVideo

Available at 2 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
Helicone
grok-4-1-fast-non-reasoning
Gateway$0.20$0.50$0.05—2,000,00030,000
QiniuGateway————2,000,0002,000,000hostTable.undisclosed

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Helicone$47.00
The cheapest paid channel is the only channel.

Price historyone sample accumulated per data sync

Input list $0.20Output list $0.50Min blended $0.28

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-14). Each future sync adds a point, and once accumulated a line is drawn here.

Related models

Grok Imagine Video 1.5 Litesame series—Grok 4.7 (Global)same series$2.00 / $6.00Grok 4.7 (US)same series$2.20 / $6.60Gemini 2.0 Flash-Litecheaper alternative$0.075 / $0.30Lyria 3 Clip Previewcheaper alternative$0 / $0Lyria 3 Pro Previewcheaper alternative$0 / $0
Data partly from models.dev (MIT) · Helicone official docs ↗