← Model list
GPT-5.4 nano
openai·openai/gpt-5.4-nano·GA·Closed·gpt-nano series
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
At a glance
Official price$0.20 / $1.25 per 1M
Cache read $0.02 · OpenAI
Lowest paid$0.18 / $1.10
Poe Gateway · 1.1× spread
Context400,000
⚠ Providers report 400,000–1,047,576; the table below is authoritative
Output limit128,000
Capabilities
✓ Reasoning✓ Tool use✓ Structured outputTemperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImagePDF
Knowledge cutoff2025-08-31
Released / updated2026-03-17 / 2026-03-17
Quality & performanceArtificial Analysis · Intelligence Index v4.1 · rep. tier xhigh
Intelligence39.7
Value97
Coding56.1
Agentic29.7
Output speed185 tok/s
TTFT4.11 s
Cost per task$0.1493
Value formulaIQ 39.7 ÷ min blended $0.410 = 97
Reasoning tier → intelligence / speed (higher tier = stronger but slower)
Non-ReasoningIQ 17.8 · 168 tok/s
mediumIQ 30.8 · 177 tok/s
xhighIQ 39.7 · 185 tok/s
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 30 providers29 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Poe | Gateway | $0.18 | $1.10 | $0.018 | — | 400,000 | 128,000 | |
| Requesty | Gateway | $0.18 | $1.13 | $0.018 | — | 400,000 | 128,000 | |
| OrcaRouter | Gateway | $0.20 | $1.25 | $0.02 | — | 400,000 | 128,000 | |
| NanoGPT | Gateway | $0.20 | $1.25 | $0.02 | — | 400,000 | 128,000 | |
| Impossibl | Gateway | $0.20 | $1.25 | $0.02 | — | 400,000 | 128,000 | |
| OpenRouter | Gateway | $0.20 | $1.25 | $0.02 | — | 400,000 | 128,000 | |
| CrossModel | Gateway | $0.20 | $1.25 | $0.02 | $0.20 | 400,000 | 128,000 | |
| Databricks databricks-gpt-5-4-nano | Cloud | $0.20 | $1.25 | $0.02 | — | 400,000 | 128,000 | |
| OpenAIOfficial | First-party | $0.20 | $1.25 | $0.02 | — | 400,000 | 128,000 | |
| Vivgrid | Gateway | $0.20 | $1.25 | $0.02 | — | 400,000 | 128,000 | |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1Poe$71.56
2Requesty$72.81
3OrcaRouter$80.90
4NanoGPT$80.90
5Impossibl$80.90
6OpenRouter$80.90
9OpenAI · Official$80.90
Switch to Poe to save $9.34/mo (12%). The gap is small, so staying on the official channel is fine.
Note: this is a gateway; verify availability and rate limits yourself.
Benchmark16 items
| Name | Conditions | Score | Metric | Source |
|---|---|---|---|---|
| SWE-Bench Pro | variant: reasoning effort xhigh | 52.4 | resolve rate | Source ↗ |
| Terminal-Bench | variant: reasoning effort xhigh · v2.0 | 46.3 | accuracy | Source ↗ |
| MCP Atlas | variant: reasoning effort xhigh | 56.1 | score | Source ↗ |
| Toolathlon | variant: reasoning effort xhigh | 35.5 | score | Source ↗ |
| τ²-Bench Telecom | variant: reasoning effort xhigh | 92.5 | accuracy | Source ↗ |
| GPQA Diamond | variant: reasoning effort xhigh | 82.8 | accuracy | Source ↗ |
| Humanity's Last Exam | variant: with tools | 37.7 | accuracy | Source ↗ |
| Humanity's Last Exam | variant: without tools | 24.3 | accuracy | Source ↗ |
| OSWorld-Verified | variant: reasoning effort xhigh | 39 | success rate | Source ↗ |
| MMMU Pro | variant: with Python | 69.5 | accuracy | Source ↗ |
| MMMU Pro | variant: without tools | 66.1 | accuracy | Source ↗ |
| OmniDocBench | variant: reasoning effort none · v1.5 | 0.2419 | overall edit distance | Source ↗ |
| OpenAI MRCR | variant: 8-needle, 64K-128K · vv2 | 44.2 | accuracy | Source ↗ |
| OpenAI MRCR | variant: 8-needle, 128K-256K · vv2 | 33.1 | accuracy | Source ↗ |
| Graphwalks | variant: BFS, 0-128K | 73.4 | accuracy | Source ↗ |
| Graphwalks | variant: parents, 0-128K | 50.8 | accuracy | Source ↗ |
The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.
Reasoning control
effort = noneeffort = loweffort = mediumeffort = higheffort = xhigheffort = maxbudget_tokenseffort = minimal
3 / 30 providers expose no reasoning control (reasoning_options: []).
Related models
gpt-5.4-nano-2026-03-17same series$0.20 / $1.25GPT-5.4 Nano (OpenAI)same series$0.20 / $1.25GPT-5.4 Nano (Azure)same series$0.20 / $1.25DeepSeek V4 Flashcheaper alternative$0.14 / $0.28GPT-5 Nanocheaper alternative$0.05 / $0.40GLM-4.7-Flashcheaper alternative$0 / $0MiMo-V2.5cheaper alternative$0.14 / $0.28
Price historyone sample accumulated per data sync
Input listOutput listMin blended