← Model list
Z-AI/GLM 4.6
zhipuai·zhipuai/glm-4.6·GA·Closed·glm series
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
At a glance
Reference price$0.45 / $1.50 per 1M
Cache read — · Helicone
Lowest paid$0.45 / $1.50
Helicone Gateway
Context204,800
⚠ Providers report 200,000–204,800; the table below is authoritative
Output limit131,072
Capabilities
Reasoning✓ Tool useStructured output✓ TemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
Text
Knowledge cutoff2024-07
Released / updated2025-10-11 / 2025-10-11
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 2 providers1 with public prices
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1Helicone$165.00
The cheapest paid channel is the only channel.
Related models
GLM Latestsame series$1.40 / $4.40Z.ai: GLM Latestsame series$1.40 / $4.40GLM 5.3 Preview Thinkingsame series$1.40 / $4.40Qwen3-Next 80B-A3B Instructcheaper alternative$0.144 / $0.574Qwen3-Coder 30B-A3B Instructcheaper alternative$0.216 / $0.861Llama-3.1-8B-Instructcheaper alternative$0.15 / $0.45Qwen3 Coder Flashcheaper alternative$0.144 / $0.574
Price historyone sample accumulated per data sync
Input listOutput listMin blended