← Model list
Llama 4 Scout
meta·meta/llama-4-scout·GA·Open weights·llama series
Open multimodal Llama model for long-context analysis and efficient agents
At a glance
Reference price$0.085 / $0.46 per 1M
Cache read $0.043 · NanoGPT
Lowest paid$0.10 / $0.30
OpenRouter Gateway · 1.2× spread
Context328,000
⚠ Providers report 328,000–1,310,720; the table below is authoritative
Output limit65,536
Capabilities
Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImage
Knowledge cutoff2024-08-31
Released / updated2025-09-05 / 2025-09-05
Quality & performanceArtificial Analysis · Intelligence Index v4.1
Intelligence10.3
Value69
Coding8.2
Agentic1.1
Output speed135 tok/s
TTFT0.78 s
Cost per task$0.0106
Value formulaIQ 10.3 ÷ min blended $0.150 = 69
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 2 providers2 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| OpenRouter | Gateway | $0.10 | $0.30 | — | — | 1,310,720 ⚠ | 16,384 | |
| NanoGPT | Gateway | $0.085 | $0.46 | $0.043 | — | 328,000 | 65,536 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1NanoGPT$34.90
2OpenRouter$35.00
The cheapest paid channel is the only channel.
Related models
Llama 3.3 70Bsame series$1.75 / $2.75Llama Guard 4 12Bsame series$0.18 / $0.18Llama 4 Mavericksame series$0.20 / $0.80Lyria 3 Clip Previewcheaper alternative$0 / $0Lyria 3 Pro Previewcheaper alternative$0 / $0
Price historyone sample accumulated per data sync
Input listOutput listMin blended