LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

GLM 5.3 Fast (Latest)

zhipuai·zhipuai/glm-fast-latest·GA·Open weights·glm series·NEW
Flagship GLM model for long-horizon coding, agents, and complex project delivery

Specs & pricing

Input / output per 1M tokens
Reference price·Fireworks AI
$2.10 / $6.60
Blended $3.23 · Cache read $0.39
Lowest paid·Fireworks AICloud
$2.10 / $6.60
Blended $3.23
Context
1,048,572
Output limit
262,144
Knowledge cutoff
—
Released / updated
2026-08-28 / 2026-09-15
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
Modalities
Text

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
Fireworks AICloud$2.10$6.60$0.39—1,048,572262,144

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Your usage cost

1Fireworks AI$544.80
The cheapest paid channel is the only channel.

Reasoning control

effort = loweffort = higheffort = max

Related models

GLM 5.3 Flash Cybersecuritysame series$0.15 / $0.50GLM Latestsame series$1.79 / $8.94GLM 5.3 Fastsame series$2.80 / $8.80DeepSeek V4 Procheaper alternative$0.435 / $0.87DeepSeek V4 Flashcheaper alternative$0.15 / $0.60GLM-5.3-Flashcheaper alternative$0.15 / $0.50DeepSeek V4.1 Flashcheaper alternative$0.15 / $0.60

Price historyone sample accumulated per data sync

Input list $2.10Output list $6.60Min blended $3.23

Price history accumulates from each data sync; currently only 1 sample(s) (2026-09-16). Each future sync adds a point, and once accumulated a line is drawn here.

Data partly from models.dev (MIT) · Fireworks AI official docs ↗