LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

GLM 5.3 Flex

zhipuai·zhipuai/glm-5.3-flex·GA·Open weights·glm series·NEW
Flagship GLM model for long-horizon coding, agents, and complex project delivery

Specs & pricing

Input / output per 1M tokens
Reference price·Neuralwatt
$0.94 / $2.93
Blended $1.44 · Cache read $0.094
Lowest paid·NeuralwattGateway
$0.94 / $2.93
Blended $1.44
Context
1,048,560
Output limit
1,048,560
Knowledge cutoff
—
Released / updated
2026-08-14 / 2026-08-14
Capabilities
✓ Reasoning✓ Tool useStructured output✓ TemperatureAttachments
Modalities
Text

Available at 1 providers1 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NeuralwattGateway$0.94$2.93$0.094—1,048,5601,048,560

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = loweffort = higheffort = maxbudget_tokens

Interleaved thinking (reasoning between tool calls) is declared by 1 of 1 providers.

Your usage cost

1Neuralwatt$232.96
The cheapest paid channel is the only channel.

Price historyone sample accumulated per data sync

Input list $0.94Output list $2.93Min blended $1.44

Price history accumulates from each data sync; currently only 1 sample(s) (2026-09-22). Each future sync adds a point, and once accumulated a line is drawn here.

Related models

GLM 5.3 Uncensoredsame series$1.25 / $2.25GLM 5.3 Primesame series$2.80 / $8.80GLM 5.3 Flash Cybersecuritysame series$0.15 / $0.50GLM-5.3-Flashcheaper alternative$0.15 / $0.50DeepSeek V4.1 Flashcheaper alternative$0.15 / $0.60DeepSeek V4 Flash 0731cheaper alternative$0.45 / $1.34GPT-5.6 Lunacheaper alternative$0.20 / $1.20
Data partly from models.dev (MIT) · Neuralwatt official docs ↗