← All models
Balanced — strong quality at a mid priceOpen Weights
GLM 5.3 Pricing
Z.ai · z-ai-glm-5-3
Input / 1M tokens
$1.40
Output / 1M tokens
$4.40
Context window
1,310,720 tokens
≈ 1,748 pages of text
Max output
943,718 tokens
What it costs in practice
Typical request (1,200 in + 400 out tokens)$0.0034
1,000 requests / month$3.44/mo10,000 requests / month$34.40/mo100,000 requests / month$344.00/moEstimate your own workloadOpens the calculator with GLM 5.3 preloaded — adjust volume and token counts there.
Price history
List price per 1M tokens since we started tracking in Aug 2026. The lines are flat because the price hasn't changed since then.
What it excels at
Supports 1M-token context and 128k output for long-document and multi-turn tasks. Low input pricing suits high-volume inference workloads.
The business tradeoff
Output token cost is elevated relative to input. Open-weights deployment introduces hosting variance and smaller ecosystem tooling.
Head-to-head
Vendor list rates, as of Sep 7, 2026 · source: openrouter · per-request examples assume 1,200 input + 400 output tokens.