LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

GLM 5.2 Fast

zhipuai·zhipuai/glm-5.2-fast·Deprecated·Open weights·glm series·NEW
Open flagship GLM for long-horizon coding agents and million-token context work

GLM 5.2 Fast by zhipuai is offered by 7 providers on this page. Its reference price is $3.00 per 1M input tokens and $10.25 per 1M output tokens. The lowest paid channel is Neuralwatt at $1.45 / $4.50 per 1M, about 2.1× below the reference price. That channel is a third-party gateway, so confirm its availability and rate limits before depending on it.

The context window is 1,048,576 tokens at the reference host, but hosts report different limits, from 1,000,000 to 1,048,576, so the usable window depends on the provider you pick. It supports reasoning, tool use, and structured output. Accepted input modalities are Text and Image. The weights are open, so it can also be self-hosted or served through a gateway of your choice. Providers report the capability flags inconsistently, so verify a specific feature against the host you plan to use.

Specs & pricing

Input / output per 1M tokens
Reference price·Wafer
$3.00 / $10.25
Blended $4.81 · Cache read $0.50
Lowest paid·NeuralwattGateway
$1.45 / $4.50
Blended $2.21 · 2.1× spread
Context
1,048,576
Output limit
131,072
Knowledge cutoff
—
Released / updated
2026-06-13 / 2026-07-13
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ TemperatureAttachments
⚠ Providers report capability flags inconsistently
Modalities
TextImage

Available at 7 providers7 with public prices

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
NeuralwattGateway$1.45$4.50$0.15—1,048,560 ⚠1,048,560deprecated
Baseten
zai-org/GLM-5.2-Fast
Cloud$2.10$6.60$0.21—1,048,576262,144
RequestyGateway$2.10$6.60$0.21—1,000,000 ⚠131,072
LLM GatewayGateway$2.20$6.50$0.45—1,000,000 ⚠131,072
above.devGateway$2.31$7.26$0.23—1,000,000 ⚠131,072
Vercel AI GatewayCloud$2.80$8.80$0.56—1,000,000 ⚠128,000
Wafer
glm5.2-fast
Gateway$3.00$10.25$0.50$01,048,576131,072

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = noneeffort = higheffort = maxbudget_tokenseffort = loweffort = mediumToggle (on / off)

Interleaved thinking (reasoning between tool calls) is declared by 4 of 7 providers.

1 / 7 providers expose no reasoning control (reasoning_options: []).

Your usage cost

1Neuralwatt$358.40
2Baseten$523.20
3Requesty$523.20
4LLM Gateway$555.00
5above.dev$575.52
6Vercel AI Gateway$731.20
The cheapest paid channel is the only channel.

Price historyone sample accumulated per data sync

Input listOutput listMin blended
$10.25$02026-08-052026-08-132026-08-05 · Input list $3.002026-08-06 · Input list $3.002026-08-07 · Input list $3.002026-08-08 · Input list $3.002026-08-09 · Input list $3.002026-08-10 · Input list $3.002026-08-11 · Input list $3.002026-08-12 · Input list $3.002026-08-13 · Input list $3.002026-08-05 · Output list $10.252026-08-06 · Output list $10.252026-08-07 · Output list $10.252026-08-08 · Output list $10.252026-08-09 · Output list $10.252026-08-10 · Output list $10.252026-08-11 · Output list $10.252026-08-12 · Output list $10.252026-08-13 · Output list $10.252026-08-05 · Min blended $2.212026-08-06 · Min blended $2.212026-08-07 · Min blended $2.212026-08-08 · Min blended $2.212026-08-09 · Min blended $2.212026-08-10 · Min blended $2.212026-08-11 · Min blended $2.212026-08-12 · Min blended $2.212026-08-13 · Min blended $2.21

Related models

GLM 5.3 Uncensoredsame series$1.25 / $2.25GLM 5.3 Primesame series$2.80 / $8.80GLM 5.3 Flash Cybersecuritysame series$0.15 / $0.50GLM-5.2cheaper alternative$1.40 / $4.40GLM-5.3cheaper alternative$1.40 / $4.40GLM-5.3-Flashcheaper alternative$0.15 / $0.50DeepSeek V4.1 Flashcheaper alternative$0.15 / $0.60
Data partly from models.dev (MIT) · Wafer official docs ↗