← Model list
Llama 4 Maverick 17B 128E Instruct
meta·meta/llama-4-maverick-17b-128e-instruct·GA·Open weights·llama series
Open multimodal Llama model for strong reasoning and fast responses
At a glance
Reference price$0.35 / $1.15 per 1M
Cache read — · Vertex
Lowest paid$0.15 / $0.60
IO.NET Cloud · 2.3× spread
1 more $0 channels
Context524,288
⚠ Providers report 128,000–524,288; the table below is authoritative
Output limit8,192
Capabilities
Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImage
Knowledge cutoff2024-02
Released / updated2025-04-01 / 2025-04-29
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 3 providers2 with public prices · 1 free
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1IO.NET$51.00
2Vertex$127.50
1 more channels offer $0 (Nvidia); free tiers usually have rate limits and no SLA, excluded from ranking.
The cheapest paid channel is the only channel.
Related models
Llama 3.3 70Bsame series$1.75 / $2.75Llama Guard 4 12Bsame series$0.18 / $0.18Llama 4 Mavericksame series$0.20 / $0.80Qwen3 Coder Flashcheaper alternative$0.144 / $0.574DeepSeek Chatcheaper alternative$0.14 / $0.28Ling-2.6-flashcheaper alternative$0.10 / $0.30Lyria 3 Clip Previewcheaper alternative$0 / $0
Price historyone sample accumulated per data sync
Input listOutput listMin blended