LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Meta-Llama-3_3-70B-Instruct

meta·meta/meta-llama-3-3-70b-instruct·GA·Open weights·llama series
Open Llama instruction model for multilingual chat, reasoning, and coding

Specs & pricing

Input / output per 1M tokens
Reference price·OVHcloud AI Endpoints
$0.74 / $0.74
Blended $0.74 · Cache read —
Lowest paid·HeliconeGateway
$0.13 / $0.39
Blended $0.20 · 5.7× spread
Context
131,072
Output limit
131,072
Knowledge cutoff
2024-12
Released / updated
2025-04-01 / 2025-04-01
Capabilities
Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
Text

Available at 2 providers2 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
Helicone
llama-3.3-70b-instruct
Gateway$0.13$0.39——128,000 ⚠16,400
OVHcloud AI Endpoints
meta-llama-3_3-70b-instruct
Cloud$0.74$0.74——131,072131,072

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Helicone$45.50
2OVHcloud AI Endpoints$185.00
The cheapest paid channel is the only channel.

Price historyone sample accumulated per data sync

Input list $0.74Output list $0.74Min blended $0.20

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-13). Each future sync adds a point, and once accumulated a line is drawn here.

Related models

Llama 3.3 70Bsame series$1.75 / $2.75Llama 4 Mavericksame series$0.25 / $0.87Llama Guard 4 12Bsame series$0.21 / $0.21Command Rcheaper alternative$0.15 / $0.60Gemma 3 12B ITcheaper alternative$0.27 / $0.27Qwen3 30B A3B Instruct 2507cheaper alternative$0.30 / $0.50DeepSeek V3.2 Expcheaper alternative$0.29 / $0.43
Data partly from models.dev (MIT) · OVHcloud AI Endpoints official docs ↗