LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

GPT-5 Mini

openai·openai/gpt-5-mini·BETA·Closed·gpt-mini series
Small GPT-5 for responsive agents, coding help, and everyday automation

GPT-5 Mini by openai is offered by 35 providers on this page. Public prices are shown for 33 of them. Its official list price is $0.25 per 1M input tokens and $2.00 per 1M output tokens. The lowest paid channel is QiHang at $0.04 / $0.29 per 1M, about 7.0× below the list price. That channel is a third-party gateway, so confirm its availability and rate limits before depending on it.

Artificial Analysis rates it 20.6 on the Intelligence Index. Against its lowest blended price of $0.102 per 1M, that is roughly 201 index points per dollar, which is the value ratio the leaderboards rank on. Median output speed is 136 tokens per second, with 10.56s to the first token.

The context window is 400,000 tokens at the reference host, but hosts report different limits, from 128,000 to 400,000, so the usable window depends on the provider you pick. It supports reasoning, tool use, and structured output. Accepted input modalities are Text, Image, and PDF. Providers report the capability flags inconsistently, so verify a specific feature against the host you plan to use. Its training knowledge cuts off at 2024-05-30.

Specs & pricing

Input / output per 1M tokens
Official price·OpenAI
$0.25 / $2.00
Blended $0.69 · Cache read $0.025
Lowest paid·QiHangGateway
$0.04 / $0.29
Blended $0.10 · 7× spread
Context
400,000
Output limit
128,000
Knowledge cutoff
2024-05-30
Released / updated
2025-08-07 / 2025-08-07
Capabilities
✓ Reasoning✓ Tool use✓ Structured outputTemperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImagePDF

Quality & performanceArtificial Analysis · Intelligence Index v4.3 · rep. tier Medium

Intelligence20.6
Value201
Coding—
Agentic—
Output speed136 tok/s
TTFT10.56 s
Value formulaIQ 20.6 ÷ min blended $0.102 = 201
Reasoning tier → intelligence / speed (higher tier = stronger but slower)
MinimalIQ 9.9 · 134 tok/s
HighIQ 16.8 · 135 tok/s
MediumIQ 20.6 · 136 tok/s

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 35 providers33 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
QiHangGateway$0.04$0.29——200,000 ⚠64,000
OfoxGateway$0.20$1.60$0.024—256,000 ⚠32,768
PoeGateway$0.22$1.80$0.022—400,000128,000
Jiekou.AIGateway$0.23$1.80——400,000128,000
Perplexity AgentGateway$0.25$2.00$0.025—400,000128,000
NanoGPTGateway$0.25$2.00$0.025—400,000128,000
NEAR AI CloudGateway$0.25$2.00$0.025—400,000128,000
Databricks
databricks-gpt-5-mini
Cloud$0.25$2.00$0.025—400,000128,000
Kilo GatewayGateway$0.25$2.00$0.025—400,000128,000
AbacusGateway$0.25$2.00$0.025—400,000128,000
OpenAIOfficialFirst-party$0.25$2.00$0.025—400,000128,000

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = minimaleffort = loweffort = mediumeffort = higheffort = xhigheffort = maxeffort = none

Interleaved thinking (reasoning between tool calls) is declared by 1 of 35 providers.

3 / 35 providers expose no reasoning control (reasoning_options: []).

Experimental modes4 items

ModeProviderInputOutputCache readCache write
fastVercel AI Gateway$0.45$3.60$0.045—
flexVercel AI Gateway$0.13$1.00$0.013—
fastOpenAI$0.45$3.60$0.045—
flexOpenAI$0.13$1.00$0.013—

Your usage cost

1QiHang$22.50
2Ofox$98.88
3Poe$110.24
4Perplexity Agent$123.00
5NanoGPT$123.00
6NEAR AI Cloud$123.00
23OpenAI · Official$123.00
Switch to QiHang to save $100.50/mo (82%)
Note: this is a gateway; verify availability and rate limits yourself.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$2.00$02026-08-052026-08-132026-08-05 · Input list $0.252026-08-06 · Input list $0.252026-08-07 · Input list $0.252026-08-08 · Input list $0.252026-08-09 · Input list $0.252026-08-10 · Input list $0.252026-08-11 · Input list $0.252026-08-12 · Input list $0.252026-08-13 · Input list $0.252026-08-05 · Output list $2.002026-08-06 · Output list $2.002026-08-07 · Output list $2.002026-08-08 · Output list $2.002026-08-09 · Output list $2.002026-08-10 · Output list $2.002026-08-11 · Output list $2.002026-08-12 · Output list $2.002026-08-13 · Output list $2.002026-08-05 · Min blended $0.102026-08-06 · Min blended $0.102026-08-07 · Min blended $0.102026-08-08 · Min blended $0.102026-08-09 · Min blended $0.102026-08-10 · Min blended $0.102026-08-11 · Min blended $0.102026-08-12 · Min blended $0.102026-08-13 · Min blended $0.10

Artificial Analysis evaluations9 items

GPQA Diamond80.3%
Humanity's Last Exam15.9%
MMLU-Pro82.8%
LiveCodeBench69.2%
Terminal-Bench Hard28.8%
AIME 202585.0%
τ²-Bench Telecom71.0%
AA-LCR72.3%
IFBench71.2%

Individual evaluations run by Artificial Analysis, on the same reasoning tier as the intelligence score above. Each benchmark has its own task set and harness, so rows are not comparable with one another. The Intelligence Index above draws on a different, newer set of evaluations.

Related models

GPT Mini Latestsame series$0.75 / $4.50GPT-5.4 minisame series$0.75 / $4.50Agentic Chat (GPT-5.4 Mini)same series$0 / $0GLM-5.3-Flashcheaper alternative$0.15 / $0.50DeepSeek V4.1 Flashcheaper alternative$0.15 / $0.60GPT-5 Nanocheaper alternative$0.05 / $0.40Qwen3.8 Flashcheaper alternative$0.15 / $0.47
Data partly from models.dev (MIT) · OpenAI official docs ↗