← Model list
Gemma 4 26B A4B
google·google/gemma-4-26b-a4b-it-2·GA·Open weights·gemma series·NEW
Open Gemma instruction model for efficient chat and self-hosted deployments
At a glance
Reference price$0.13 / $0.40 per 1M
Cache read — · NovitaAI
Lowest paid$0.05 / $0.29
EmpirioLabs AI Gateway · 2.6× spread
Context262,144
Output limit131,072
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
⚠ Providers report capability flags inconsistently
Modalities
TextImageVideo
Knowledge cutoff—
Released / updated2026-04-02 / 2026-06-12
Quality & performanceArtificial Analysis · Intelligence Index v4.1 · rep. tier Reasoning
Intelligence26.1
Value237
Coding39.3
Agentic11
Output speed—
TTFT—
Cost per task$0.0391
Value formulaIQ 26.1 ÷ min blended $0.110 = 237
Reasoning tier → intelligence / speed (higher tier = stronger but slower)
Non-reasoningIQ 20.4 · 75 tok/s
ReasoningIQ 26.1
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
Available at 4 providers4 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| EmpirioLabs AI gemma-4-26b-a4b | Gateway | $0.05 | $0.29 | $0.025 | — | 262,144 | 32,768 | |
| NanoGPT google/gemma-4-26b-a4b-it | Gateway | $0.13 | $0.40 | $0.065 | — | 262,144 | 131,072 | |
| Merge Gateway google/gemma-4-26b-a4b-it | Gateway | $0.13 | $0.40 | — | — | 262,144 | 65,536 | |
| NovitaAI google/gemma-4-26b-a4b-it | Gateway | $0.13 | $0.40 | — | — | 262,144 | 131,072 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1EmpirioLabs AI$21.50
2NanoGPT$38.20
3Merge Gateway$46.00
4NovitaAI$46.00
The cheapest paid channel is the only channel.
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
Toggle (on / off)effort = noneeffort = loweffort = mediumeffort = higheffort = maxbudget_tokens ≥ 128 ≤ 32,768
Related models
Gemma 4 12B Instructsame series$0.06 / $0.30Google Gemma 3 27B Instructsame series$0.12 / $0.20Google Gemma 4 26B A4B Instructsame series$0.13 / $0.40GLM-4.7-Flashcheaper alternative$0 / $0Hy3cheaper alternative$0 / $0Nemotron 3 Nano 30B A3Bcheaper alternative$0 / $0Qwen3.7 Flashcheaper alternative$0.03 / $0.118
Price historyone sample accumulated per data sync
Input listOutput listMin blended