GLM 5.3 FlashX
Current indexed sources identify z-ai/glm-5.3-flashx as Z.ai GLM 5.3 FlashX. $0.37 per 1M input tokens, $1.25 per 1M output tokens, and $0.075 per 1M cached-input read tokens; a 1,048,576-token context window; a 131,072-token maximum completion; September 18, 2026 as the release date; text/image/video input with text output; tool calling; and JSON response-format support. Direct official Z.ai documentation and pricing pages were not located in the current indexed sources, so fields sourced through OpenRouter should be validated before staging approval.
- Text
- Image
- Video
- Text
Pricing
Every billing leg from the live catalog, in USD per 1M tokens.
| Billing leg | Rate |
|---|---|
| Input | $0.37 / M |
| Output | $1.25 / M |
| Cache read | $0.075 / M |
Cost calculator
Billing-grade math over the rate card above, including the thinking tokens most estimates miss.
Provider
Where this model runs, and what it charges.
| Platform | Input / M | Output / M | Cache read / M | Batch | Access |
|---|---|---|---|---|---|
Z.ai | $0.37 | $1.25 | $0.075 | - | via OpenRouter |
Price history
Has this model gotten cheaper or more expensive?
More from Z.ai
Other models from the same provider.
Stop estimating. Track it.
SuperPenguin meters every GLM 5.3 FlashX call your team makes, prices it with this exact rate card, and shows spend by feature, customer, and team.
Rates are generated from SuperPenguin's live pricing catalog.