computetrail

All models ·z-ai

Z.ai: GLM 5.3

Provider: z-ai · z-ai/glm-5.3

Input
$1.4 /1M
Output
$4.4 /1M
Cache read
$0.26 /1M
Context
1,048,576 tokens

How Z.ai: GLM 5.3 pricing compares

At $4.4 / 1M output tokens, Z.ai: GLM 5.3 is the 14th-cheapest of 14 z-ai models we track. It's cheaper than 32% of the 392 paid models on computetrail (266th-cheapest overall). Output costs 3.1× its input price ($1.4 / 1M in). For a typical 3:1 input-to-output workload that works out to about $2.15 / 1M blended. Reading cached input costs 81% less than fresh input, so prompt caching cuts repeat-context cost sharply. Its 1,048,576 tokens context window ranks 34th-largest of 392 models with a published limit.

Output price history

$4.40$4.4008-1908-2308-27

Across 9 daily snapshots since 2026-08-19, the output price has held steady at $4.4 / 1M — no change so far.

Similar-priced models

ModelProviderOutput $/1MContext
Qwen: Qwen3.7 Maxqwen$4.4251,000,000
Google: Gemini 3.5 Flash (batch)google$4.51,048,576
MoonshotAI: Kimi K2.7 Code (batch)moonshotai$4262,144
MoonshotAI: Kimi K2.6moonshotai$4262,144
OpenAI: o3 (batch)openai$4200,000
Z.ai: GLM 5.1z-ai$3.96204,800

Prices are recorded daily from public provider data via OpenRouter and normalized to USD per 1M tokens. See methodology.