← Model list
Llama-3.2-3B
meta·meta/llama-3.2-3b·GA·Open weights·llama series
Small open Llama base model for lightweight text generation and self-hosting
At a glance
Reference price$0.15 / $0.60 per 1M
Cache read — · Venice AI
Lowest paid$0.10 / $0.10
Pioneer Gateway · 1.5× spread
Context128,000
⚠ Providers report 128,000–131,072; the table below is authoritative
Output limit4,096
Capabilities
ReasoningTool use? Structured output✓ TemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
Text
Knowledge cutoff2023-12
Released / updated2024-09-25 / 2024-09-25
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 2 providers2 with public prices
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1Pioneer$25.00
2Venice AI$60.00
The cheapest paid channel is the only channel.
Related models
Llama 3.3 70Bsame series$1.75 / $2.75Llama Guard 4 12Bsame series$0.18 / $0.18Llama 4 Mavericksame series$0.20 / $0.80Mistral Nemocheaper alternative$0.15 / $0.15Llama 3.2 1B Instructcheaper alternative$0.10 / $0.201Llama 4 Maverick 17B 128E Instruct FP8cheaper alternative$0 / $0Mistral Nemo Instruct 2407cheaper alternative$0.145 / $0.145
Price historyone sample accumulated per data sync
Input listOutput listMin blended