LLM Pricing
PricingLeaderboardsToolsProvidersReleasesGuides

© 2026 LLM Pricing

About
·Contact
·Privacy
·RSS
← Model list

GLM-5.3-Flash (free)

zhipuai·zhipuai/glm-5.3-flash-free·GA·Open weights·glm-flash series·NEW
Native multimodal GLM model for efficient coding and long-horizon agent tasks

Specs & pricing

Input / output per 1M tokens
Reference price·OrcaRouter
$0 / $0
Blended $0 · Cache read —
Lowest paid
No paid channels
1 more $0 channels
Context
1,000,000
Output limit
128,000
Knowledge cutoff
—
Released / updated
2026-08-26 / 2026-08-26
Capabilities
✓ Reasoning✓ Tool use✓ Structured output✓ Temperature✓ Attachments
Modalities
TextImageVideoPDF

Available at 1 providers0 with public prices · 1 free

ProviderTierInputOutputCache readCache writeContextOutput limitStatus
OrcaRouterGatewayFree——1,000,000128,000

Sorted by blended price (input×0.75 + output×0.25) asc. The official channel always shows regardless of rank. Whether a gateway's low price is actually usable can't be verified.

Reasoning control

effort = loweffort = higheffort = max

Interleaved thinking (reasoning between tool calls) is declared by 1 of 1 providers.

Your usage cost

This model is $0 across all listed channels (OrcaRouter).

Price historyone sample accumulated per data sync

Input list $0Output list $0Min blended —

Price history accumulates from each data sync; currently only 1 sample(s) (2026-09-09). Each future sync adds a point, and once accumulated a line is drawn here.

Related models

GLM-5.3-FlashXsame series$0.37 / $1.25GLM Flash (latest)same series$0.23 / $0.75GLM 5.3 Flash Uncensoredsame series$0.20 / $0.80
Data partly from models.dev (MIT) · OrcaRouter official docs ↗