LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

DiffusionGemma

google·google/diffusiongemma·GA·Open weights·NEW
DiffusionGemma is a high-speed diffusion-based version of Gemma 4 26B A4B. It supports optional reasoning and a 262,144-token context window.

Specs & pricing

Input / output per 1M tokens
Reference price·NanoGPT
$0.05 / $0.15
Blended $0.075 · Cache read $0.025
Lowest paid·NanoGPTGateway
$0.05 / $0.15
Blended $0.075
Context
262,144
Output limit
32,768
Knowledge cutoff
—
Released / updated
2026-09-19 / 2026-09-19
Capabilities
✓ ReasoningTool useStructured output? TemperatureAttachments
Modalities
Text

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NanoGPTGateway$0.05$0.15$0.025—262,14432,768

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = noneeffort = xhigh

Your usage cost

1NanoGPT$14.50
The cheapest paid channel is the only channel.

Price historyone sample accumulated per data sync

Input list $0.05Output list $0.15Min blended $0.075

Price history accumulates from each data sync; currently only 1 sample(s) (2026-09-20). Each future sync adds a point, and once accumulated a line is drawn here.

Related models

Hy3cheaper alternative$0 / $0Nemotron 3.5 Lightning 30B A3Bcheaper alternative$0 / $0Laguna S 2.1cheaper alternative$0 / $0GLM-4.5-Flashcheaper alternative$0 / $0
Data partly from models.dev (MIT) · NanoGPT official docs ↗