LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

GPT-4.1 mini

openai·openai/gpt-4.1-mini·Deprecated·Closed·gpt-mini series
Affordable GPT-4.1 lane for fast coding help and structured extraction

GPT-4.1 mini by openai is offered by 26 providers on this page. Public prices are shown for 24 of them. Its official list price is $0.40 per 1M input tokens and $1.60 per 1M output tokens. The lowest paid channel is Ofox at $0.32 / $1.28 per 1M, about 1.4× below the list price. That channel is a third-party gateway, so confirm its availability and rate limits before depending on it.

Artificial Analysis rates it 10.2 on the Intelligence Index, with 20.2 for coding. Against its lowest blended price of $0.56 per 1M, that is roughly 18 index points per dollar, which is the value ratio the leaderboards rank on. Median output speed is 173 tokens per second, with 0.72s to the first token.

The context window is 1,047,576 tokens at the reference host, but hosts report different limits, from 1,000,000 to 1,047,576, so the usable window depends on the provider you pick. It supports tool use and structured output. Accepted input modalities are Text, Image, and PDF. Providers report the capability flags inconsistently, so verify a specific feature against the host you plan to use. Its training knowledge cuts off at 2024-04.

Specs & pricing

Input / output per 1M tokens
Official price·OpenAI
$0.40 / $1.60
Blended $0.70 · Cache read $0.10
Lowest paid·OfoxGateway
$0.32 / $1.28
Blended $0.56 · 1.4× spread
Context
1,047,576
Output limit
32,768
Knowledge cutoff
2024-04
Released / updated
2025-04-14 / 2025-04-14
Capabilities
Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImagePDF

Quality & performanceArtificial Analysis · Intelligence Index v4.3

Intelligence10.2
Value18
Coding20.2
Agentic—
Output speed173 tok/s
TTFT0.72 s
Value formulaIQ 10.2 ÷ min blended $0.560 = 18

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 26 providers24 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
OfoxGateway$0.32$1.28$0.08—1,047,57632,768
PoeGateway$0.36$1.40$0.09—1,047,57632,768
NanoGPTGateway$0.40$1.60$0.10—1,047,57632,768
NEAR AI CloudGateway$0.40$1.60$0.10—1,047,57632,768
Kilo GatewayGateway$0.40$1.60$0.10—1,047,57632,768
AbacusGateway$0.40$1.60$0.10—1,047,57632,768
302.AIGateway$0.40$1.60——1,047,57632,768
OpenRouterGateway$0.40$1.60$0.10—1,047,57632,768
AzureCloud$0.40$1.60$0.10—1,047,57632,768deprecated
Azure Cognitive ServicesCloud$0.40$1.60$0.10—1,047,57632,768deprecated
OpenAIOfficialFirst-party$0.40$1.60$0.10—1,047,57632,768

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Experimental modes2 items

ModeProviderInputOutputCache readCache write
fastVercel AI Gateway$0.70$2.80$0.18—
fastOpenAI$0.70$2.80$0.18—

Your usage cost

1Ofox$99.20
2Poe$109.60
3NanoGPT$124.00
4NEAR AI Cloud$124.00
5Kilo Gateway$124.00
6Abacus$124.00
18OpenAI · Official$124.00
Switch to Ofox to save $24.80/mo (20%)
Note: this is a gateway; verify availability and rate limits yourself.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$1.60$02026-08-052026-09-182026-08-05 · Input list $0.402026-08-06 · Input list $0.402026-08-07 · Input list $0.402026-08-08 · Input list $0.402026-08-09 · Input list $0.402026-08-10 · Input list $0.402026-08-11 · Input list $0.402026-08-12 · Input list $0.402026-08-13 · Input list $0.402026-09-18 · Input list $0.402026-08-05 · Output list $1.602026-08-06 · Output list $1.602026-08-07 · Output list $1.602026-08-08 · Output list $1.602026-08-09 · Output list $1.602026-08-10 · Output list $1.602026-08-11 · Output list $1.602026-08-12 · Output list $1.602026-08-13 · Output list $1.602026-09-18 · Output list $1.602026-08-05 · Min blended $0.622026-08-06 · Min blended $0.622026-08-07 · Min blended $0.622026-08-08 · Min blended $0.622026-08-09 · Min blended $0.622026-08-10 · Min blended $0.622026-08-11 · Min blended $0.622026-08-12 · Min blended $0.622026-08-13 · Min blended $0.622026-09-18 · Min blended $0.56

Benchmark1 items

NameConditionsScoreMetricSource
Aider Polyglot—32.4percent correctSource ↗

The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.

Artificial Analysis evaluations13 items

GPQA Diamond66.4%
Humanity's Last Exam5.0%
MMLU-Pro78.1%
LiveCodeBench48.3%
Terminal-Bench Hard7.6%
Terminal-Bench 2.110.1%
AIME 202546.3%
AIME43.0%
MATH-50092.5%
τ²-Bench Telecom52.9%
τ³-Bench Banking5.4%
AA-LCR44.0%
IFBench38.3%

Individual evaluations run by Artificial Analysis, on the same reasoning tier as the intelligence score above. Each benchmark has its own task set and harness, so rows are not comparable with one another. The Intelligence Index above draws on a different, newer set of evaluations.

Related models

GPT Mini Latestsame series$0.75 / $4.50GPT-5.4 minisame series$0.75 / $4.50Agentic Chat (GPT-5.4 Mini)same series$0 / $0Grok 4.1 Fast (Non-Reasoning)cheaper alternative$0.20 / $0.50grok-4-fast-non-reasoningcheaper alternative$0.20 / $0.50Gemini 2.0 Flash-Litecheaper alternative$0.075 / $0.30Lyria 3 Clip Previewcheaper alternative$0 / $0
Data partly from models.dev (MIT) · OpenAI official docs ↗