← All models
Flash — fastest and cheapest, lighter reasoningCommercial API
Gemini 3.1 Flash Lite Pricing
Google · google-gemini-3-1-flash-lite
Input / 1M tokens
$0.250
Output / 1M tokens
$1.50
Context window
1,048,576 tokens
≈ 1,398 pages of text
Max output
65,536 tokens
What it costs in practice
Typical request (1,200 in + 400 out tokens)$0.0009
1,000 requests / month$0.9/mo10,000 requests / month$9.00/mo100,000 requests / month$90.00/moEstimate your own workloadOpens the calculator with Gemini 3.1 Flash Lite preloaded — adjust volume and token counts there.
Price history
Only one price point recorded so far.
The staircase chart appears once a price change is detected.
What it excels at
Handles high-volume inference and long-context multimodal inputs at low cost. Supports 1M-token context with competitive throughput for classification, extraction, and summarization workloads.
The business tradeoff
Reduced reasoning depth and instruction-following accuracy relative to larger Gemini variants. Output token pricing remains higher than input, which can increase costs on generative tasks.
Cheaper from Google
Head-to-head
Vendor list rates, as of Jun 15, 2026 · source: openrouter · per-request examples assume 1,200 input + 400 output tokens.