Efficient GLM model for fast reasoning, coding, and agent workflows
GLM-4.5-Flash by zhipuai is offered by 5 providers on this page. Its official list price is $0 per 1M tokens. It also has 5 free ($0) channels; free tiers usually carry rate limits, and subscription-covered access bills $0 per token only after the subscription fee.
The context window is 131,072 tokens at the reference host, but hosts report different limits, from 131,072 to 200,000, so the usable window depends on the provider you pick. It supports reasoning, tool use, and structured output. Providers report the capability flags inconsistently, so verify a specific feature against the host you plan to use. Its training knowledge cuts off at 2025-04.
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Z.AIOfficial | First-party | Free | $0 | $0 | 131,072 | 98,304 | ||
| EmpirioLabs AI glm-4-5-flash | Gateway | Free | — | — | 200,000 ⚠ | 98,304 | ||
| UnoRouter glm-4.5-flash:free | Gateway | Free | — | — | 131,072 | 98,304 | ||
| Tempr Gateway | Gateway | Free | $0 | $0 | 131,072 | 98,304 | ||
| Zhipu AIOfficial | First-party | Free | $0 | $0 | 131,072 | 98,304 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
1 / 5 providers expose no reasoning control (reasoning_options: []).
This model is $0 across all listed channels (Z.AI, EmpirioLabs AI, UnoRouter, Tempr Gateway, Zhipu AI).