← Model list
GLM-4.5-Flash
zhipuai·zhipuai/glm-4.5-flash·GA·Closed·glm-flash series
Efficient GLM model for fast reasoning, coding, and agent workflows
At a glance
Official price$0 / $0 per 1M
Cache read $0 · Z.AI
Lowest paidNo paid channels
4 more $0 channels
Context131,072
⚠ Providers report 131,072–200,000; the table below is authoritative
Output limit98,304
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
Text
Knowledge cutoff2025-04
Released / updated2025-07-28 / 2025-07-28
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 4 providers0 with public prices · 4 free
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| Z.AIOfficial | First-party | Free | $0 | $0 | 131,072 | 98,304 | ||
| UnoRouter glm-4.5-flash:free | Gateway | Free | — | — | 131,072 | 98,304 | ||
| Zhipu AIOfficial | First-party | Free | $0 | $0 | 131,072 | 98,304 | ||
| EmpirioLabs AI glm-4-5-flash | Gateway | Free | — | — | 200,000 ⚠ | 98,304 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
This model is $0 across all listed channels (Z.AI, UnoRouter, Zhipu AI, EmpirioLabs AI).
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
Toggle (on / off)
1 / 4 providers expose no reasoning control (reasoning_options: []).
Related models
GLM-4.7-Flashsame series$0 / $0GLM 4.7 Flash Thinkingsame series$0.07 / $0.40GLM-4.7-Flash (Free)same series$0 / $0
Price historyone sample accumulated per data sync
Input listOutput listMin blended