← Model list
GPT-4o (Fast)
openai·openai/gpt-4o-fast·GA·Closed·gpt series
Omni-era GPT for multimodal chat, practical coding, and general assistants
At a glance
Reference price$4.25 / $17.00 per 1M
Cache read $2.13 · Vercel AI Gateway
Lowest paid$4.25 / $17.00
Vercel AI Gateway Cloud
Context128,000
Output limit16,384
Capabilities
Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImagePDF
Knowledge cutoff2023-09
Released / updated2024-05-13 / 2024-08-06
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 1 providers1 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Vercel AI Gateway | Cloud | $4.25 | $17.00 | $2.13 | — | 128,000 | 16,384 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1Vercel AI Gateway$1,445.00
The cheapest paid channel is the only channel.
Related models
RouteLLMsame series$3.00 / $15.00GPT-Realtime-2.1same series$4.00 / $24.00NanoGPT Helpsame series$0 / $0Qwen3 Maxcheaper alternative$0.861 / $3.44Qwen3 235B-A22B Instruct 2507cheaper alternative$1.03 / $3.08Qwen3-Next 80B-A3B Instructcheaper alternative$0.144 / $0.574Qwen3-Coder 30B-A3B Instructcheaper alternative$0.216 / $0.861
Price historyone sample accumulated per data sync
Input list $4.25Output list $17.00Min blended $7.44
Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-19). Each future sync adds a point, and once accumulated a line is drawn here.