Z-AI

Z.ai: GLM 5.3 FlashX

z-ai/glm-5.3-flashx-20260918

GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...

toolsstructuredreasoningimage in
Data source

Z.ai: GLM 5.3 FlashX usage over time

OpenRouter share of tokens, by day.

View
Tidelines — LLM usage and market share

Z.ai: GLM 5.3 FlashX holds 0.34% today, up from 0.34% at the start of the visible window.

Data: OpenRouter as of 2026-09-22 · Source: tidelines.ai/models/z-ai-glm-5-3-flashx-20260918

Rank history

Tidelines — LLM usage and market share
Data: OpenRouter as of 2026-09-22 · Source: tidelines.ai/models/z-ai-glm-5-3-flashx-20260918

Adoption curve vs the pantheon

Z.ai: GLM 5.3 FlashX's share plotted against days since first tracked, overlaid on other current leaders on the same axis. The question every new-model page asks: is this one going to make it?

Tidelines — LLM usage and market share
Data: OpenRouter as of 2026-09-22 · Source: tidelines.ai/models/z-ai-glm-5-3-flashx-20260918

Citable stat

As of 2026-09-22, Z.ai: GLM 5.3 FlashX holds 0.3% of OpenRouter token share, up 0.00pp over 30 days. Source: Tidelines (OpenRouter).

Where to run it

Z.ai: GLM 5.3 FlashX pricing by provider

Tidelines — LLM usage and market share

Per-provider pricing, context, and uptime via OpenRouter. Sorted by output price.

Z.ai: GLM 5.3 FlashX is served by 1 tracked endpoint on OpenRouter. The cheapest output price is $1.25 per million tokens from Z.AI. The largest context window offered is 1049k tokens. Prices are list rates and refresh daily.

ProviderInput $/MOutput $/MContextUptime
Z.AI$0.37$1.251049k100.00%