Efficient GLM model for fast reasoning, coding, and agent workflows
GLM-4.7-FlashX by zhipuai is offered by 9 providers on this page. Its official list price is $0.07 per 1M input tokens and $0.40 per 1M output tokens. The lowest paid channel is Vercel AI Gateway at $0.06 / $0.40 per 1M, about 1.2× below the list price.
The context window is 200,000 tokens, with an output limit of 131,072 tokens. It supports reasoning and tool use. The weights are open, so it can also be self-hosted or served through a gateway of your choice. Its training knowledge cuts off at 2025-04.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Vercel AI Gateway | Cloud | $0.06 | $0.40 | $0.01 | — | 200,000 | 128,000 | |
| Z.AIOfficial | First-party | $0.07 | $0.40 | $0.01 | $0 | 200,000 | 131,072 | |
| Merge Gateway | Gateway | $0.07 | $0.40 | $0.01 | $0 | 200,000 | 131,072 | |
| Tempr Gateway | Gateway | $0.07 | $0.40 | $0.01 | $0 | 200,000 | 131,072 | |
| Zhipu AIOfficial | First-party | $0.07 | $0.40 | $0.01 | $0 | 200,000 | 131,072 | |
| DevPass (LLM Gateway) | Gateway | $0.07 | $0.40 | $0.01 | $0 | 200,000 | 131,072 | |
| LLM Gateway | Gateway | $0.07 | $0.40 | $0.01 | — | 200,000 | 128,000 | |
| Ofox | Gateway | $0.072 | $0.40 | $0.01 | — | 200,000 | 128,000 | |
| ZenMux | Gateway | $0.073 | $0.44 | $0.015 | — | 200,000 | 128,000 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Interleaved thinking (reasoning between tool calls) is declared by 2 of 9 providers.