Z-AI

Z.ai: GLM 5.3 Flash

z-ai/glm-5.3-flash-20260826Hugging Face ↗

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

toolsstructuredreasoningimage in
Data source

Z.ai: GLM 5.3 Flash usage over time

OpenRouter share of tokens, by day.

View
Tidelines — LLM usage and market share

Z.ai: GLM 5.3 Flash holds 11.57% today, up from 1.96% at the start of the visible window.

Data: OpenRouter as of 2026-09-01 · Source: tidelines.ai/models/z-ai-glm-5-3-flash-20260826

Rank history

Tidelines — LLM usage and market share
Data: OpenRouter as of 2026-09-01 · Source: tidelines.ai/models/z-ai-glm-5-3-flash-20260826

Adoption curve vs the pantheon

Z.ai: GLM 5.3 Flash's share plotted against days since first tracked, overlaid on other current leaders on the same axis. The question every new-model page asks: is this one going to make it?

Tidelines — LLM usage and market share
Data: OpenRouter as of 2026-09-01 · Source: tidelines.ai/models/z-ai-glm-5-3-flash-20260826

Citable stat

As of 2026-09-01, Z.ai: GLM 5.3 Flash holds 11.6% of OpenRouter token share, up 9.61pp over 30 days. Source: Tidelines (OpenRouter).

Cost per session

Median dollars spent on one Z.ai: GLM 5.3 Flash session, by coding agent and session length.

Tidelines — LLM usage and market share

Token rate sets the price; session length and reasoning set the bill.

Agent1 turn2–9 turns10–49 turns50+ turns
Claude Code$0.0022$0.0052$0.038$0.343
Kilo Code$0.0011$0.0059$0.030$0.255
Hermes Agent$0.0015$0.0040$0.028$0.205
Codex$0.0009$0.037

Compare every model's session cost →

Data: OpenRouter session-cost dataset as of 2026-08-30 · Source: tidelines.ai/models/z-ai-glm-5-3-flash-20260826

Where to run it

Z.ai: GLM 5.3 Flash pricing by provider

Tidelines — LLM usage and market share

Per-provider pricing, context, and uptime via OpenRouter. Sorted by output price.

Z.ai: GLM 5.3 Flash is served by 23 tracked endpoints on OpenRouter. The cheapest output price is $0.25 per million tokens from GMICloud, and the most expensive is $0.50 from Cloudflare. The largest context window offered is 1311k tokens. Prices are list rates and refresh daily.

ProviderInput $/MOutput $/MContextUptime
GMICloud$0.07$0.251049k97.35%
Novita$0.07$0.251049k99.19%
DeepInfra$0.07$0.251049k99.49%
Z.AI$0.07$0.251049k97.13%
Wafer$0.10$0.351049k98.27%
Morph$0.13$0.451049k99.41%
Makora$0.14$0.471049k98.14%
Modal$0.15$0.501049k99.93%
Sail Research$0.15$0.501049k99.75%
StreamLake$0.15$0.501024k98.85%
NextBit$0.15$0.501049k97.75%
Fireworks$0.15$0.501049k99.93%
Phala$0.15$0.501049k98.12%
Friendli$0.15$0.501049k99.69%
SiliconFlow$0.15$0.501049k95.09%
DigitalOcean$0.15$0.501049k98.74%
Together$0.15$0.501049k99.05%
Reka$0.15$0.50262k99.92%
Parasail$0.15$0.501049k94.19%
BaseTen$0.15$0.501049k99.30%
Venice$0.15$0.501049k99.02%
Io Net$0.15$0.50262k86.19%
Cloudflare$0.15$0.501311k99.98%