LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

OpenAI GPT-4.1 Mini

openai·openai/gpt-4.1-mini-2025-04-14·GA·Closed·gpt-mini series
Compact GPT model for low-latency assistance and high-volume workloads

Specs & pricing

Input / output per 1M tokens
Reference price·Helicone
$0.40 / $1.60
Blended $0.70 · Cache read $0.10
Lowest paid·HeliconeGateway
$0.40 / $1.60
Blended $0.70 · 1× spread
Context
1,047,576
Output limit
32,768
Knowledge cutoff
2025-04
Released / updated
2025-04-14 / 2025-04-14
Capabilities
Reasoning✓ Tool use? Structured output✓ TemperatureAttachments
Modalities
TextImage

Available at 2 providers2 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
HeliconeGateway$0.40$1.60$0.10—1,047,57632,768
Helicone
gpt-4.1-mini
Gateway$0.40$1.60$0.10—1,047,57632,768

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Helicone$124.00
2Helicone$124.00
The cheapest paid channel is the only channel.

Price historyone sample accumulated per data sync

Input list $0.40Output list $1.60Min blended $0.70

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-13). Each future sync adds a point, and once accumulated a line is drawn here.

Related models

GPT Mini Latestsame series$0.75 / $4.50GPT-5.4 minisame series$0.75 / $4.50Agentic Chat (GPT-5.4 Mini)same series$0 / $0Grok 4.1 Fast (Non-Reasoning)cheaper alternative$0.20 / $0.50grok-4-fast-non-reasoningcheaper alternative$0.20 / $0.50Gemini 2.0 Flash-Litecheaper alternative$0.075 / $0.30Lyria 3 Clip Previewcheaper alternative$0 / $0
Data partly from models.dev (MIT) · Helicone official docs ↗