LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

o4-mini

openai·openai/o4-mini·Deprecated·Closed·o-mini series
Fast o-series model for compact reasoning, coding, and tool use

o4-mini by openai is offered by 21 providers on this page. Public prices are shown for 19 of them. Its official list price is $1.10 per 1M input tokens and $4.40 per 1M output tokens.

Artificial Analysis rates it 16.7 on the Intelligence Index. Against its lowest blended price of $1.74 per 1M, that is roughly 10 index points per dollar, which is the value ratio the leaderboards rank on. Median output speed is 152 tokens per second, with 25.87s to the first token.

The context window is 200,000 tokens, with an output limit of 100,000 tokens. It supports reasoning, tool use, and structured output. Accepted input modalities are Text, Image, and PDF. Providers report the capability flags inconsistently, so verify a specific feature against the host you plan to use. Its training knowledge cuts off at 2024-05.

Specs & pricing

Input / output per 1M tokens
Official price·OpenAI
$1.10 / $4.40
Blended $1.93 · Cache read $0.28
Lowest paid·PoeGateway
$0.99 / $4.00
Blended $1.74 · 1.1× spread
Context
200,000
Output limit
100,000
Knowledge cutoff
2024-05
Released / updated
2025-04-16 / 2025-04-16
Capabilities
✓ Reasoning✓ Tool use✓ Structured outputTemperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImagePDF

Quality & performanceArtificial Analysis · Intelligence Index v4.3 · rep. tier High

Intelligence16.7
Value10
Coding—
Agentic—
Output speed152 tok/s
TTFT25.87 s
Value formulaIQ 16.7 ÷ min blended $1.742 = 10

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 21 providers19 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
PoeGateway$0.99$4.00$0.25—200,000100,000
NEAR AI CloudGateway$1.10$4.40$0.28—200,000100,000
Kilo GatewayGateway$1.10$4.40$0.28—200,000100,000
AbacusGateway$1.10$4.40——200,000100,000
OpenRouterGateway$1.10$4.40$0.28—200,000100,000
AzureCloud$1.10$4.40$0.28—200,000100,000deprecated
Azure Cognitive ServicesCloud$1.10$4.40$0.28—200,000100,000deprecated
ImpossiblGateway$1.10$4.40$0.28—200,000100,000
Cloudflare AI GatewayCloud$1.10$4.40$0.28—200,000100,000
Jiekou.AIGateway$1.10$4.40——200,000100,000
OpenAIOfficialFirst-party$1.10$4.40$0.28—200,000100,000deprecated

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = loweffort = mediumeffort = higheffort = noneToggle (on / off)effort = xhigh

1 / 21 providers expose no reasoning control (reasoning_options: []).

Experimental modes4 items

ModeProviderInputOutputCache readCache write
fastVercel AI Gateway$2.00$8.00$0.50—
flexVercel AI Gateway$0.55$2.20$0.14—
fastOpenAI$2.00$8.00$0.50—
flexOpenAI$0.55$2.20$0.14—

Your usage cost

1Poe$309.20
2NEAR AI Cloud$341.00
3Kilo Gateway$341.00
4OpenRouter$341.00
5Azure$341.00
6Azure Cognitive Services$341.00
14OpenAI · Official$341.00
Switch to Poe to save $31.80/mo (9%). The gap is small, so staying on the official channel is fine.
Note: this is a gateway; verify availability and rate limits yourself.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$4.40$02026-08-052026-09-032026-08-05 · Input list $1.102026-08-06 · Input list $1.102026-08-07 · Input list $1.102026-08-08 · Input list $1.102026-08-09 · Input list $1.102026-08-10 · Input list $1.102026-08-11 · Input list $1.102026-08-12 · Input list $1.102026-08-13 · Input list $1.102026-09-03 · Input list $1.102026-08-05 · Output list $4.402026-08-06 · Output list $4.402026-08-07 · Output list $4.402026-08-08 · Output list $4.402026-08-09 · Output list $4.402026-08-10 · Output list $4.402026-08-11 · Output list $4.402026-08-12 · Output list $4.402026-08-13 · Output list $4.402026-09-03 · Output list $4.402026-08-05 · Min blended $1.742026-08-06 · Min blended $1.742026-08-07 · Min blended $1.742026-08-08 · Min blended $1.742026-08-09 · Min blended $1.742026-08-10 · Min blended $1.742026-08-11 · Min blended $1.742026-08-12 · Min blended $1.742026-08-13 · Min blended $1.932026-09-03 · Min blended $1.74

Benchmark1 items

NameConditionsScoreMetricSource
Aider Polyglot—72percent correctSource ↗

The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.

Artificial Analysis evaluations11 items

GPQA Diamond78.4%
Humanity's Last Exam16.5%
MMLU-Pro83.2%
LiveCodeBench85.9%
Terminal-Bench Hard15.2%
AIME 202590.7%
AIME94.0%
MATH-50098.9%
τ²-Bench Telecom55.6%
AA-LCR61.0%
IFBench68.7%

Individual evaluations run by Artificial Analysis, on the same reasoning tier as the intelligence score above. Each benchmark has its own task set and harness, so rows are not comparable with one another. The Intelligence Index above draws on a different, newer set of evaluations.

Related models

GPT Audio Minisame series$0.60 / $2.40o4-mini (EU)same series$1.21 / $4.84o4-mini (Fast)same series$2.00 / $8.00GLM-5.3-Flashcheaper alternative$0.15 / $0.50DeepSeek V4.1 Flashcheaper alternative$0.15 / $0.60DeepSeek V4 Flash 0731cheaper alternative$0.45 / $1.34GPT-5.6 Lunacheaper alternative$0.20 / $1.20
Data partly from models.dev (MIT) · OpenAI official docs ↗