Z.ai

GLM-5.2

Z.ai's latest flagship GLM model for agentic coding and reasoning, available first-party and across inference gateways.

ActiveAvailable from 31 providers
Tool callingCan invoke external tools and APIs (function calling) mid-response.Structured outputCan return JSON or schema-constrained responses instead of free text.ReasoningUses extended thinking / chain-of-thought for harder multi-step problems.Effort controlLets you dial thinking depth (low / medium / high) to trade cost for quality.Fine-tuningSupports training a custom variant on your own data (vendor fine-tune API or self-hosted open weights).

Source: Z.ai GLM API pricing, Z.ai GLM models

Price per 1M tokens via Novita (OpenRouter)
$0.2744 in / $0.8624 out
Context window204,800
Max output131,072
Cache read$0.05096 / M
ModalitiesText → Text
Priced across31 providers

Pricing

Every billing leg from the live catalog, in USD per 1M tokens.

Billing legRate
Input$0.2744 / M
Output$0.8624 / M
Cache read$0.05096 / M

Cost calculator

Billing-grade math over the rate card above, including the thinking tokens most estimates miss.

Hidden reasoning tokens bill as output. 1.0x = thinking off. Medium-effort reasoning typically lands near 3x; heavy reasoning runs 6x to 8x. Set it to match your workload.

Estimated monthly cost
$107
at current list prices
Fresh input$54.88
Cached input$0.00
Output incl. thinking$51.74
Track this automatically →

Compare providers

The same model, priced across every platform that serves it. Lowest combined in/out rate is flagged.

PlatformInput / MOutput / MCache read / MBatchAccess
NovitaLowest
$0.2744$0.8624$0.05096via OpenRouter
StreamLake
$0.2751$0.8646$0.05109via OpenRouter
Baidu
$0.28$0.88$0.052via OpenRouter
Inceptron
$0.94$2.9$0.17via OpenRouter
DeepInfra
$0.93$3$0.18via OpenRouter
GMICloud
$0.98$3.08$0.182via OpenRouter
Alibaba
$1.24$3.89$0.24724via OpenRouter
Morph
$1.1$4.1$0.22via OpenRouter
AtlasCloud
$1.26$3.96$0.234via OpenRouter
Wafer
$1.2$4.1$0.2via OpenRouter
SiliconFlow
$1.3$4.09$0.26via OpenRouter
Decart
$1.2$4.2$0.2via OpenRouter
Ambient
$1.05$4.4$0.2via OpenRouter
DigitalOcean
$1.05$4.4$0.21via OpenRouter
Akashml
$1.3$4.4via OpenRouter
Wandb
$1.39$4.4$0.26via OpenRouter
Together AI
$1.4$4.4$0.26Direct
Cursor
$1.4$4.4$0.26Direct
Fireworks AI
$1.4$4.4$0.14Direct
Baseten
$1.4$4.4$0.26via OpenRouter
Chutes
$1.4$4.4$0.7via OpenRouter
Cloudflare
$1.4$4.4$0.26via OpenRouter
Fireworks AI
$1.4$4.4$0.14via OpenRouter
Friendli
$1.4$4.4$0.26via OpenRouter
Ionstream
$1.4$4.4$0.26via OpenRouter
Parasail
$1.4$4.4$0.26via OpenRouter
Phala
$1.4$4.4$0.7via OpenRouter
Together AI
$1.4$4.4$0.26via OpenRouter
Venice
$1.4$4.4$0.26via OpenRouter
Z.ai
$1.4$4.4$0.26via OpenRouter
Io-net
$1.6$4.99$0.7986via OpenRouter

Price history

Has this model gotten cheaper or more expensive?

EffectiveInput / OutputStatus
July 17, 2026 – July 19, 2026$0.9324 / $2.93past
July 19, 2026$0.2744 / $0.8624current

More from Z.ai

Other models from the same provider.

Stop estimating. Track it.

SuperPenguin meters every GLM-5.2 call your team makes, prices it with this exact rate card, and shows spend by feature, customer, and team.

Get started free

Rates are generated from SuperPenguin's live pricing catalog.