LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

GPT-4.1 nano

openai·openai/gpt-4.1-nano·Deprecated·Closed·gpt-nano series
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks

GPT-4.1 nano by openai is offered by 23 providers on this page. Its official list price is $0.10 per 1M input tokens and $0.40 per 1M output tokens. The lowest paid channel is SAP AI Core at $0.08 / $0.26 per 1M, about 1.4× below the list price.

Artificial Analysis rates it 9.6 on the Intelligence Index, with 11.1 for coding and 1.2 for agentic tasks. Against its lowest blended price of $0.125 per 1M, that is roughly 77 index points per dollar, which is the value ratio the leaderboards rank on. Median output speed is 149 tokens per second, with 0.7s to the first token. Running one task of the Artificial Analysis suite costs about $0.0349, which reflects how many tokens its reasoning consumes rather than the unit price alone.

The context window is 1,047,576 tokens at the reference host, but hosts report different limits, from 1,000,000 to 1,047,576, so the usable window depends on the provider you pick. It supports tool use and structured output. Accepted input modalities are Text, Image, and PDF. Providers report the capability flags inconsistently, so verify a specific feature against the host you plan to use. Its training knowledge cuts off at 2024-04.

Specs & pricing

Input / output per 1M tokens
Official price·OpenAI
$0.10 / $0.40
Blended $0.175 · Cache read $0.025
Lowest paid·SAP AI CoreCloud
$0.08 / $0.26
Blended $0.125 · 1.4× spread
Context
1,047,576
Output limit
32,768
Knowledge cutoff
2024-04
Released / updated
2025-04-14 / 2025-04-14
Capabilities
Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImagePDF

Quality & performanceArtificial Analysis · Intelligence Index v4.1

Intelligence9.6
Value77
Coding11.1
Agentic1.2
Output speed149 tok/s
TTFT0.7 s
Cost per task$0.0349
Value formulaIQ 9.6 ÷ min blended $0.125 = 77

Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.

Available at 23 providers23 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
SAP AI CoreCloud$0.08$0.26——1,047,57632,768
PoeGateway$0.09$0.36$0.022—1,047,57632,768
HeliconeGateway$0.10$0.40$0.025—1,047,57632,768
NanoGPTGateway$0.10$0.40$0.025—1,047,57632,768
ImpossiblGateway$0.10$0.40$0.025—1,047,57632,768
OrcaRouterGateway$0.10$0.40$0.025—1,047,57632,768
Azure Cognitive ServicesFirst-party$0.10$0.40$0.025—1,047,57632,768deprecated
PioneerGateway$0.10$0.40$0.05$0.101,047,57632,768
Vercel AI GatewayCloud$0.10$0.40$0.025—1,047,57632,768deprecated
DevPass (LLM Gateway)Gateway$0.10$0.40$0.025—1,000,000 ⚠32,768
OpenAIOfficialFirst-party$0.10$0.40$0.025—1,047,57632,768deprecated

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Poe$27.84
2SAP AI Core$29.00
3Helicone$31.00
4NanoGPT$31.00
5Impossibl$31.00
6OrcaRouter$31.00
17OpenAI · Official$31.00
Switch to Poe to save $3.16/mo (10%). The gap is small, so staying on the official channel is fine.
Note: this is a gateway; verify availability and rate limits yourself.

Benchmark1 items

NameConditionsScoreMetricSource
Aider Polyglot—8.9percent correctSource ↗

The same benchmark scores very differently across harness / dataset, so the qualifying conditions must be shown together.

Related models

gpt-5.4-nano-2026-03-17same series$0.20 / $1.25GPT-5.4 nanosame series$0.20 / $1.25Agentic Chat (GPT-5.4 Nano)same series$0 / $0Lyria 3 Clip Previewcheaper alternative$0 / $0Lyria 3 Pro Previewcheaper alternative$0 / $0

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$0.40$02026-08-052026-08-132026-08-05 · Input list $0.102026-08-06 · Input list $0.102026-08-07 · Input list $0.102026-08-08 · Input list $0.102026-08-09 · Input list $0.102026-08-10 · Input list $0.102026-08-11 · Input list $0.102026-08-12 · Input list $0.102026-08-13 · Input list $0.102026-08-05 · Output list $0.402026-08-06 · Output list $0.402026-08-07 · Output list $0.402026-08-08 · Output list $0.402026-08-09 · Output list $0.402026-08-10 · Output list $0.402026-08-11 · Output list $0.402026-08-12 · Output list $0.402026-08-13 · Output list $0.402026-08-05 · Min blended $0.1252026-08-06 · Min blended $0.1252026-08-07 · Min blended $0.1252026-08-08 · Min blended $0.1252026-08-09 · Min blended $0.1252026-08-10 · Min blended $0.1252026-08-11 · Min blended $0.1252026-08-12 · Min blended $0.1252026-08-13 · Min blended $0.125
Data partly from models.dev (MIT) · OpenAI official docs ↗