← Model list
GLM 4.7 Flash Original Thinking
zhipuai·zhipuai/glm-4.7-flash-original-thinking·GA·Open weights·glm series
Efficient GLM model for fast reasoning, coding, and agent workflows
At a glance
Reference price$0.07 / $0.40 per 1M
Cache read $0.035 · NanoGPT
Lowest paid$0.07 / $0.40
NanoGPT Gateway
Context200,000
Output limit128,000
Capabilities
✓ Reasoning✓ Tool useStructured output? TemperatureAttachments
Modalities
Text
Knowledge cutoff—
Released / updated2026-01-19 / 2026-01-19
Quality & performance
Artificial Analysis doesn't cover this model (267 of 2059 have data). Quality data comes from independent evals covering widely used models.
Available at 1 providers1 with public prices
| Provider | Tier | Input | Output | Cache read | Cache write | Context | Output limit | Status |
|---|---|---|---|---|---|---|---|---|
| NanoGPT zai-org/glm-4.7-flash-original:thinking | Gateway | $0.07 | $0.40 | $0.035 | — | 200,000 | 128,000 |
Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.
Your usage cost
1NanoGPT$29.80
The cheapest paid channel is the only channel.
Benchmark
No upstream benchmark data for this model. For quality, see the Artificial Analysis intelligence score above.
Reasoning control
Upstream provides no control info
1 / 1 providers expose no reasoning control (reasoning_options: []).
Related models
GLM Latestsame series$1.40 / $4.40Z.ai: GLM Latestsame series$1.40 / $4.40GLM 5.3 Preview Thinkingsame series$1.40 / $4.40GLM-4.7-Flashcheaper alternative$0 / $0Hy3cheaper alternative$0 / $0Nemotron 3 Nano 30B A3Bcheaper alternative$0 / $0Qwen3.7 Flashcheaper alternative$0.03 / $0.118
Price historyone sample accumulated per data sync
Input listOutput listMin blended