← All models
Flash — fastest and cheapest, lighter reasoningOpen Weights
GLM 5.3 (batch) Pricing
Z.ai · z-ai-glm-5-3-batch
Input / 1M tokens
$0.700
Output / 1M tokens
$2.20
Context window
1,048,576 tokens
≈ 1,398 pages of text
Max output
943,718 tokens
What it costs in practice
Typical request (1,200 in + 400 out tokens)$0.0017
1,000 requests / month$1.72/mo10,000 requests / month$17.20/mo100,000 requests / month$172.00/moEstimate your own workloadOpens the calculator with GLM 5.3 (batch) preloaded — adjust volume and token counts there.
Price history
Only one price point recorded so far.
The staircase chart appears once a price change is detected.
What it excels at
Supports 1M-token contexts with high output limits, suited for large-scale batch text workloads. Open weights enable local or custom deployment.
The business tradeoff
Performance and latency vary by hosting provider. Smaller ecosystem and fewer integrated tools than closed flagship models.
Head-to-head
Vendor list rates, as of Sep 8, 2026 · source: openrouter · per-request examples assume 1,200 input + 400 output tokens.