LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

Gemma 4 31B IT FP8

google·google/gemma-4-31b-it-fp8·GA·Open weights·gemma series
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning

Specs & pricing

Input / output per 1M tokens
Reference price·InferX
$0 / $0
Blended $0 · Cache read —
Lowest paid
No paid channels
1 more $0 channels
Context
262,144
Output limit
32,768
Knowledge cutoff
—
Released / updated
2026-04-02 / 2026-04-02
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImage

Available at 1 providers0 with public prices · 1 free

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
InferX
gemma-4-31B-it-fp8
GatewayFree——262,14432,768

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

Upstream provides no control info

1 / 1 providers expose no reasoning control (reasoning_options: []).

Your usage cost

This model is $0 across all listed channels (InferX).

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$0.0001$02026-08-052026-08-132026-08-05 · Input list $02026-08-06 · Input list $02026-08-07 · Input list $02026-08-08 · Input list $02026-08-09 · Input list $02026-08-10 · Input list $02026-08-11 · Input list $02026-08-12 · Input list $02026-08-13 · Input list $02026-08-05 · Output list $02026-08-06 · Output list $02026-08-07 · Output list $02026-08-08 · Output list $02026-08-09 · Output list $02026-08-10 · Output list $02026-08-11 · Output list $02026-08-12 · Output list $02026-08-13 · Output list $0

Related models

Gemma 4 26B A4B Cybersecuritysame series$0.11 / $0.33Gemma 4 31B Split-Untiedsame series$0.10 / $0.30Gemma 4 12B Semancersame series$0.05 / $0.25
Data partly from models.dev (MIT) · InferX official docs ↗