← Model list
Gemma 3N E4B Instruct
google·google/gemma-3n-e4b-it-3·GA·Open weights·gemma series
Open Gemma instruction model for efficient chat and self-hosted deployments
At a glance
Reference price$0.06 / $0.12 per 1M
Cache read — · Together AI
Lowest paid$0.06 / $0.12
Together AI Cloud
Context32,768
Output limit32,768
Capabilities
ReasoningTool use✓ Structured output✓ TemperatureAttachments
Modalities
Text
Knowledge cutoff—
Released / updated2025-05-20 / 2025-05-20
Quality & performanceArtificial Analysis · Intelligence Index v4.1
Intelligence1
Value13
Coding3.2
Agentic—
Output speed55 tok/s
TTFT1.5 s
Value formulaIQ 1 ÷ min blended $0.075 = 13
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 1 providers1 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Together AI google/gemma-3n-E4B-it | Cloud | $0.06 | $0.12 | — | — | 32,768 | 32,768 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1Together AI$18.00
The cheapest paid channel is the only channel.
Related models
Gemma 4 12B Instructsame series$0.06 / $0.30Gemma 4 26B A4Bsame series$0.13 / $0.40Google Gemma 3 27B Instructsame series$0.12 / $0.20Llama 4 Maverick 17B 128E Instruct FP8cheaper alternative$0 / $0Lyria 3 Clip Previewcheaper alternative$0 / $0Lyria 3 Pro Previewcheaper alternative$0 / $0
Price historyone sample accumulated per data sync
Input listOutput listMin blended