Flagship GLM model for long-horizon coding, agents, and complex project delivery
GLM 5.3 Thinking by zhipuai is currently listed from a single provider. Its reference price is $1.00 per 1M input tokens and $3.20 per 1M output tokens.
The context window is 1,048,576 tokens, with an output limit of 131,072 tokens. It supports reasoning, tool use, and structured output. The weights are open, so it can also be self-hosted or served through a gateway of your choice.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NanoGPT z-ai/glm-5.3:thinking | Gateway | $1.00 | $3.20 | $0.20 | — | 1,048,576 | 131,072 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.