chooseaimodel
← All models
Flashfastest and cheapest, lighter reasoningCommercial API

Gemma 3 4B Pricing

Google · google-gemma-3-4b-it

ShareXFacebookLinkedIn
Input / 1M tokens
$0.050
Output / 1M tokens
$0.100
Context window
131,072 tokens
175 pages of text
Max output
16,384 tokens

What it costs in practice

Estimate your own workloadOpens the calculator with Gemma 3 4B preloaded — adjust volume and token counts there.

Price history

Only one price point recorded so far.

The staircase chart appears once a price change is detected.

What it excels at

Low-cost, low-latency inference for basic text and vision tasks with 128k context support.

The business tradeoff

Weaker complex reasoning and instruction following than models above 20B parameters.

Same tier from other providers

Head-to-head

Vendor list rates, as of Jun 15, 2026 · source: openrouter · per-request examples assume 1,200 input + 400 output tokens.