← Model list
GPT-5.4 mini
openai·openai/gpt-5.4-mini·GA·Closed·gpt-mini series
Strong small GPT for coding subagents, quick tool use, and high-volume work
At a glance
Official price$0.75 / $4.50 per 1M
Cache read $0.075 · OpenAI
Lowest paid$0.375 / $4.00
Xpersona Gateway · 2.5× spread
1 more $0 channels
Context400,000
⚠ Providers report 272,000–400,000; the table below is authoritative
Output limit128,000
Capabilities
✓ Reasoning✓ Tool use✓ Structured outputTemperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImagePDF
Knowledge cutoff2025-08-31
Released / updated2026-03-17 / 2026-03-17
Quality & performanceArtificial Analysis · Intelligence Index v4.1 · rep. tier xhigh
Intelligence40.9
Value32
Coding56.1
Agentic31.5
Output speed163 tok/s
TTFT11.89 s
Cost per task$0.4948
Value formulaIQ 40.9 ÷ min blended $1.281 = 32
Reasoning tier → intelligence / speed (higher tier = stronger but slower)
Non-ReasoningIQ 16.8 · 162 tok/s
mediumIQ 30.5 · 160 tok/s
xhighIQ 40.9 · 163 tok/s
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 36 providers34 with public prices · 1 free
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Kenari gpt-5-4-mini | Gateway | Free | — | — | 400,000 | 128,000 | ||
| Xpersona | Gateway | $0.375 | $4.00 | $0.037 | — | 272,000 ⚠ | 128,000 | |
| Poe | Gateway | $0.68 | $4.00 | $0.068 | — | 400,000 | 128,000 | |
| Requesty | Gateway | $0.675 | $4.05 | $0.068 | — | 400,000 | 128,000 | |
| OrcaRouter | Gateway | $0.75 | $4.50 | $0.075 | — | 400,000 | 128,000 | |
| NanoGPT | Gateway | $0.75 | $4.50 | $0.075 | — | 400,000 | 128,000 | |
| Impossibl | Gateway | $0.75 | $4.50 | $0.075 | — | 400,000 | 128,000 | |
| OpenRouter | Gateway | $0.75 | $4.50 | $0.075 | — | 400,000 | 128,000 | |
| CrossModel | Gateway | $0.75 | $4.50 | $0.075 | $0.75 | 400,000 | 128,000 | |
| Databricks databricks-gpt-5-4-mini | Cloud | $0.75 | $4.50 | $0.075 | — | 400,000 | 128,000 | |
| OpenAIOfficial | First-party | $0.75 | $4.50 | $0.075 | — | 400,000 | 128,000 | |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1Xpersona$234.50
2Poe$262.56
3Requesty$264.60
4OrcaRouter$294.00
5NanoGPT$294.00
6Impossibl$294.00
10OpenAI · Official$294.00
1 more channels offer $0 (Kenari); free tiers usually have rate limits and no SLA, excluded from ranking.
Switch to Xpersona to save $59.50/mo (20%)
Note: this is a gateway; verify availability and rate limits yourself.
Benchmark16 items
| Name | Conditions | Score | Metric | Source |
|---|---|---|---|---|
| SWE-Bench Pro | variant: reasoning effort xhigh | 54.4 | resolve rate | Source ↗ |
| Terminal-Bench | variant: reasoning effort xhigh · v2.0 | 60 | accuracy | Source ↗ |
| MCP Atlas | variant: reasoning effort xhigh | 57.7 | score | Source ↗ |
| Toolathlon | variant: reasoning effort xhigh | 42.9 | score | Source ↗ |
| τ²-Bench Telecom | variant: reasoning effort xhigh | 93.4 | accuracy | Source ↗ |
| GPQA Diamond | variant: reasoning effort xhigh | 88 | accuracy | Source ↗ |
| Humanity's Last Exam | variant: with tools | 41.5 | accuracy | Source ↗ |
| Humanity's Last Exam | variant: without tools | 28.2 | accuracy | Source ↗ |
| OSWorld-Verified | variant: reasoning effort xhigh | 72.1 | success rate | Source ↗ |
| MMMU Pro | variant: with Python | 78 | accuracy | Source ↗ |
| MMMU Pro | variant: without tools | 76.6 | accuracy | Source ↗ |
| OmniDocBench | variant: reasoning effort none · v1.5 | 0.1263 | overall edit distance | Source ↗ |
| OpenAI MRCR | variant: 8-needle, 64K-128K · vv2 | 47.7 | accuracy | Source ↗ |
| OpenAI MRCR | variant: 8-needle, 128K-256K · vv2 | 33.6 | accuracy | Source ↗ |
| Graphwalks | variant: BFS, 0-128K | 76.3 | accuracy | Source ↗ |
| Graphwalks | variant: parents, 0-128K | 71.5 | accuracy | Source ↗ |
The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.
Reasoning control
effort = noneeffort = loweffort = mediumeffort = higheffort = xhigheffort = maxbudget_tokenseffort = minimal
2 / 36 providers expose no reasoning control (reasoning_options: []).
Experimental modes5 items
| Mode | Provider | Input | Output | Cache read | Cache write |
|---|---|---|---|---|---|
| fast | OrcaRouter | $1.50 | $9.00 | $0.15 | — |
| fast | Databricks | $1.50 | $9.00 | $0.15 | — |
| fast | OpenAI | $1.50 | $9.00 | $0.15 | — |
| fast | Merge Gateway | $1.50 | $9.00 | $0.15 | — |
| fast | AIHubMix | $1.50 | $9.00 | $0.15 | — |
Related models
OpenAI GPT Mini Latestsame series$0.75 / $4.50gpt-5.4-mini-2026-03-17same series$0.75 / $4.50GPT-5.4 Mini (OpenAI)same series$0.75 / $4.50DeepSeek V4 Procheaper alternative$0.435 / $0.87DeepSeek V4 Flashcheaper alternative$0.14 / $0.28MiniMax-M2.7cheaper alternative$0.30 / $1.20DeepSeek V4 Flash 0731cheaper alternative$0.479 / $1.44
Price historyone sample accumulated per data sync
Input listOutput listMin blended