← Model list

Meta-Llama-3_3-70B-Instruct

meta·meta/meta-llama-3-3-70b-instruct·GA·Open weights·llama series

Open Llama instruction model for multilingual chat, reasoning, and coding

At a glance
Reference price$0.74 / $0.74 per 1M
Cache read · OVHcloud AI Endpoints
Lowest paid$0.13 / $0.39
Helicone Gateway · 5.7× spread
Context131,072
⚠ Providers report 128,000–131,072; the table below is authoritative
Output limit131,072
Capabilities
ReasoningTool useStructured outputTemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
Text
Knowledge cutoff2024-12
Released / updated2025-04-01 / 2025-04-01

Quality & performance

Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.

Available at 2 providers2 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
Helicone
llama-3.3-70b-instruct
Gateway$0.13$0.39128,00016,400
OVHcloud AI Endpoints
meta-llama-3_3-70b-instruct
Cloud$0.74$0.74131,072131,072

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Helicone$45.50
2OVHcloud AI Endpoints$185.00
The cheapest paid channel is the only channel.

Related models

Price historyone sample accumulated per data sync

Input list $0.74Output list $0.74Min blended $0.195

Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-13). Each future sync adds a point, and once accumulated a line is drawn here.