Z.ai: GLM 5.3
Provider: z-ai · z-ai/glm-5.3
How Z.ai: GLM 5.3 pricing compares
At $4.4 / 1M output tokens, Z.ai: GLM 5.3 is the 14th-cheapest of 14 z-ai models we track. It's cheaper than 32% of the 392 paid models on computetrail (266th-cheapest overall). Output costs 3.1× its input price ($1.4 / 1M in). For a typical 3:1 input-to-output workload that works out to about $2.15 / 1M blended. Reading cached input costs 81% less than fresh input, so prompt caching cuts repeat-context cost sharply. Its 1,048,576 tokens context window ranks 34th-largest of 392 models with a published limit.
Output price history
Across 9 daily snapshots since 2026-08-19, the output price has held steady at $4.4 / 1M — no change so far.
Similar-priced models
| Model | Provider | Output $/1M | Context |
|---|---|---|---|
| Qwen: Qwen3.7 Max | qwen | $4.425 | 1,000,000 |
| Google: Gemini 3.5 Flash (batch) | $4.5 | 1,048,576 | |
| MoonshotAI: Kimi K2.7 Code (batch) | moonshotai | $4 | 262,144 |
| MoonshotAI: Kimi K2.6 | moonshotai | $4 | 262,144 |
| OpenAI: o3 (batch) | openai | $4 | 200,000 |
| Z.ai: GLM 5.1 | z-ai | $3.96 | 204,800 |
Prices are recorded daily from public provider data via OpenRouter and normalized to USD per 1M tokens. See methodology.