Z.ai

GLM 5.3 FlashX

Current indexed sources identify z-ai/glm-5.3-flashx as Z.ai GLM 5.3 FlashX. $0.37 per 1M input tokens, $1.25 per 1M output tokens, and $0.075 per 1M cached-input read tokens; a 1,048,576-token context window; a 131,072-token maximum completion; September 18, 2026 as the release date; text/image/video input with text output; tool calling; and JSON response-format support. Direct official Z.ai documentation and pricing pages were not located in the current indexed sources, so fields sourced through OpenRouter should be validated before staging approval.

Active
Tool callingCan invoke external tools and APIs (function calling) mid-response.Structured outputCan return JSON or schema-constrained responses instead of free text.ReasoningUses extended thinking / chain-of-thought for harder multi-step problems.Effort controlLets you dial thinking depth (low / medium / high) to trade cost for quality.Fine-tuningSupports training a custom variant on your own data (vendor fine-tune API or self-hosted open weights).

Source: https://openrouter.ai/z-ai/glm-5.3-flashx

Price per 1M tokens via Z.ai (OpenRouter)
$0.37 in / $1.25 out
Context window1,048,576
Max output131,072
Cache read$0.075 / M
Input
  • Text
  • Image
  • Video
Output
  • Text

Pricing

Every billing leg from the live catalog, in USD per 1M tokens.

Billing legRate
Input$0.37 / M
Output$1.25 / M
Cache read$0.075 / M

Cost calculator

Billing-grade math over the rate card above, including the thinking tokens most estimates miss.

Estimated monthly cost
$99.00
at current list prices
Fresh input$74.00
Cached input$0.00
Output$25.00
Track this automatically →

Provider

Where this model runs, and what it charges.

PlatformInput / MOutput / MCache read / MBatchAccess
Z.ai
$0.37$1.25$0.075-via OpenRouter

Price history

Has this model gotten cheaper or more expensive?

No changePriced at $0.37 / $1.25 per 1M since September 19, 2026.

More from Z.ai

Other models from the same provider.

Stop estimating. Track it.

SuperPenguin meters every GLM 5.3 FlashX call your team makes, prices it with this exact rate card, and shows spend by feature, customer, and team.

Get started free

Rates are generated from SuperPenguin's live pricing catalog.