chooseaimodel
← All models
Balancedstrong quality at a mid priceOpen Weights

Llama 3.1 70B Instruct Pricing

Meta · meta-llama-llama-3-1-70b-instruct

ShareXFacebookLinkedIn
Input / 1M tokens
$0.400
Output / 1M tokens
$0.400
Context window
131,072 tokens
175 pages of text
Max output
16,384 tokens

What it costs in practice

Estimate your own workloadOpens the calculator with Llama 3.1 70B Instruct preloaded — adjust volume and token counts there.

Price history

List price per 1M tokens since we started tracking in Jun 2026. Each step marks a price change — hover a point for the exact date and rate.

What it excels at

Strong performance on instruction following, coding, and multilingual tasks with 128k context support. Competitive results on standard benchmarks at low per-token cost.

The business tradeoff

Quality and latency depend on hosting provider; lacks native tool-use and vision capabilities present in some closed models of similar size.

Cheaper from Meta

Same tier from other providers

Head-to-head

Vendor list rates, as of Sep 12, 2026 · source: openrouter · per-request examples assume 1,200 input + 400 output tokens.