← All models
Balanced — strong quality at a mid priceOpen Weights
Llama 3.1 70B Instruct Pricing
Meta · meta-llama-llama-3-1-70b-instruct
Input / 1M tokens
$0.400
Output / 1M tokens
$0.400
Context window
131,072 tokens
≈ 175 pages of text
Max output
16,384 tokens
What it costs in practice
Typical request (1,200 in + 400 out tokens)$0.0006
1,000 requests / month$0.64/mo10,000 requests / month$6.40/mo100,000 requests / month$64.00/moEstimate your own workloadOpens the calculator with Llama 3.1 70B Instruct preloaded — adjust volume and token counts there.
Price history
List price per 1M tokens since we started tracking in Jun 2026. Each step marks a price change — hover a point for the exact date and rate.
What it excels at
Strong performance on instruction following, coding, and multilingual tasks with 128k context support. Competitive results on standard benchmarks at low per-token cost.
The business tradeoff
Quality and latency depend on hosting provider; lacks native tool-use and vision capabilities present in some closed models of similar size.
Cheaper from Meta
Head-to-head
Vendor list rates, as of Sep 12, 2026 · source: openrouter · per-request examples assume 1,200 input + 400 output tokens.