Native multimodal GLM model for efficient coding and long-horizon agent tasks
Quality is independently evaluated by Artificial Analysis. Speed/latency are model-level medians.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Zhipu AI Coding Plan | Gateway | Subscription | $0 | $0 | 1,000,000 | 131,072 | ||
| Z.AI Coding Plan | Gateway | Subscription | $0 | $0 | 1,000,000 | 131,072 | ||
| NanoGPT | Gateway | $0.075 | $0.25 | $0.015 | — | 1,048,576 ⚠ | 131,072 | |
| OpenRouter | Gateway | $0.075 | $0.25 | $0.015 | — | 1,310,720 ⚠ | 131,072 | |
| Z.AIOfficial | First-party | $0.075 | $0.25 | $0.015 | $0 | 1,000,000 | 131,072 | |
| Merge Gateway | Gateway | $0.075 | $0.25 | $0.015 | — | 1,000,000 | 131,072 | |
| Zhipu AIOfficial | First-party | $0.075 | $0.25 | $0.015 | $0 | 1,000,000 | 131,072 | |
| EmpirioLabs AI glm-5-3-flash | Gateway | $0.075 | $0.25 | $0.075 | — | 1,000,000 | 131,072 | |
| Kilo Gateway | Gateway | $0.075 | $0.25 | $0.015 | — | 1,048,576 ⚠ | 131,072 | |
| Venice AI z-ai-glm-5-3-flash | Gateway | $0.094 | $0.313 | $0.019 | — | 1,048,576 ⚠ | 131,072 | |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
2 more channels offer $0 (Zhipu AI Coding Plan, Z.AI Coding Plan); free tiers usually have rate limits and no SLA, excluded from ranking.
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
2 / 18 providers expose no reasoning control (reasoning_options: []).
Price history accumulates from each data sync; currently only 1 sample(s) (2026-08-27). Each future sync adds a point, and once accumulated a line is drawn here.