← All models
Flash — fastest and cheapest, lighter reasoningOpen Weights
GLM 4.7 Flash Pricing
Z.ai · z-ai-glm-4-7-flash
Input / 1M tokens
$0.060
Output / 1M tokens
$0.400
Context window
202,752 tokens
≈ 270 pages of text
Max output
16,384 tokens
What it costs in practice
Typical request (1,200 in + 400 out tokens)$0.0002
1,000 requests / month$0.232/mo10,000 requests / month$2.32/mo100,000 requests / month$23.20/moEstimate your own workloadOpens the calculator with GLM 4.7 Flash preloaded — adjust volume and token counts there.
Price history
List price per 1M tokens since we started tracking in Jun 2026. Each step marks a price change — hover a point for the exact date and rate.
What it excels at
Handles large-context text tasks at very low cost. Open weights support local or private deployment.
The business tradeoff
Reasoning depth trails flagship closed models. Output quality and consistency vary with hosting implementation.
Head-to-head
Vendor list rates, as of Jul 23, 2026 · source: openrouter · per-request examples assume 1,200 input + 400 output tokens.