Z-AI
Z.ai: GLM 5.3 FlashX
z-ai/glm-5.3-flashx-20260918
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
Z.ai: GLM 5.3 FlashX usage over time
OpenRouter share of tokens, by day.
Z.ai: GLM 5.3 FlashX holds 0.34% today, up from 0.34% at the start of the visible window.
Rank history
Adoption curve vs the pantheon
Z.ai: GLM 5.3 FlashX's share plotted against days since first tracked, overlaid on other current leaders on the same axis. The question every new-model page asks: is this one going to make it?
Citable stat
As of 2026-09-22, Z.ai: GLM 5.3 FlashX holds 0.3% of OpenRouter token share, up 0.00pp over 30 days. Source: Tidelines (OpenRouter).
Where to run it
Z.ai: GLM 5.3 FlashX pricing by provider
Per-provider pricing, context, and uptime via OpenRouter. Sorted by output price.
Z.ai: GLM 5.3 FlashX is served by 1 tracked endpoint on OpenRouter. The cheapest output price is $1.25 per million tokens from Z.AI. The largest context window offered is 1049k tokens. Prices are list rates and refresh daily.
| Provider | Input $/M | Output $/M | Context | Uptime |
|---|---|---|---|---|
| Z.AI | $0.37 | $1.25 | 1049k | 100.00% |