GLM-5.2
Z.ai's latest flagship GLM model for agentic coding and reasoning, available first-party and across inference gateways.
Source: Z.ai GLM API pricing, Z.ai GLM models
Pricing
Every billing leg from the live catalog, in USD per 1M tokens.
| Billing leg | Rate |
|---|---|
| Input | $0.2744 / M |
| Output | $0.8624 / M |
| Cache read | $0.05096 / M |
Cost calculator
Billing-grade math over the rate card above, including the thinking tokens most estimates miss.
Hidden reasoning tokens bill as output. 1.0x = thinking off. Medium-effort reasoning typically lands near 3x; heavy reasoning runs 6x to 8x. Set it to match your workload.
Compare providers
The same model, priced across every platform that serves it. Lowest combined in/out rate is flagged.
| Platform | Input / M | Output / M | Cache read / M | Batch | Access |
|---|---|---|---|---|---|
NovitaLowest | $0.2744 | $0.8624 | $0.05096 | — | via OpenRouter |
StreamLake | $0.2751 | $0.8646 | $0.05109 | — | via OpenRouter |
Baidu | $0.28 | $0.88 | $0.052 | — | via OpenRouter |
Inceptron | $0.94 | $2.9 | $0.17 | — | via OpenRouter |
DeepInfra | $0.93 | $3 | $0.18 | — | via OpenRouter |
GMICloud | $0.98 | $3.08 | $0.182 | — | via OpenRouter |
Alibaba | $1.24 | $3.89 | $0.24724 | — | via OpenRouter |
Morph | $1.1 | $4.1 | $0.22 | — | via OpenRouter |
AtlasCloud | $1.26 | $3.96 | $0.234 | — | via OpenRouter |
Wafer | $1.2 | $4.1 | $0.2 | — | via OpenRouter |
SiliconFlow | $1.3 | $4.09 | $0.26 | — | via OpenRouter |
Decart | $1.2 | $4.2 | $0.2 | — | via OpenRouter |
Ambient | $1.05 | $4.4 | $0.2 | — | via OpenRouter |
DigitalOcean | $1.05 | $4.4 | $0.21 | — | via OpenRouter |
Akashml | $1.3 | $4.4 | — | — | via OpenRouter |
Wandb | $1.39 | $4.4 | $0.26 | — | via OpenRouter |
Together AI | $1.4 | $4.4 | $0.26 | — | Direct |
Cursor | $1.4 | $4.4 | $0.26 | — | Direct |
Fireworks AI | $1.4 | $4.4 | $0.14 | ✓ | Direct |
Baseten | $1.4 | $4.4 | $0.26 | — | via OpenRouter |
Chutes | $1.4 | $4.4 | $0.7 | — | via OpenRouter |
Cloudflare | $1.4 | $4.4 | $0.26 | — | via OpenRouter |
Fireworks AI | $1.4 | $4.4 | $0.14 | — | via OpenRouter |
Friendli | $1.4 | $4.4 | $0.26 | — | via OpenRouter |
Ionstream | $1.4 | $4.4 | $0.26 | — | via OpenRouter |
Parasail | $1.4 | $4.4 | $0.26 | — | via OpenRouter |
Phala | $1.4 | $4.4 | $0.7 | — | via OpenRouter |
Together AI | $1.4 | $4.4 | $0.26 | — | via OpenRouter |
Venice | $1.4 | $4.4 | $0.26 | — | via OpenRouter |
Z.ai | $1.4 | $4.4 | $0.26 | — | via OpenRouter |
Io-net | $1.6 | $4.99 | $0.7986 | — | via OpenRouter |
Price history
Has this model gotten cheaper or more expensive?
| Effective | Input / Output | Status |
|---|---|---|
| July 17, 2026 – July 19, 2026 | $0.9324 / $2.93 | past |
| July 19, 2026 | $0.2744 / $0.8624 | current |
More from Z.ai
Other models from the same provider.
Stop estimating. Track it.
SuperPenguin meters every GLM-5.2 call your team makes, prices it with this exact rate card, and shows spend by feature, customer, and team.
Rates are generated from SuperPenguin's live pricing catalog.






