← All models
Flash — fastest and cheapest, lighter reasoningOpen Weights
Qwen3 VL 8B Instruct Pricing
Qwen · qwen-qwen3-vl-8b-instruct
Input / 1M tokens
$0.117
Output / 1M tokens
$0.455
Context window
262,144 tokens
≈ 350 pages of text
Max output
32,768 tokens
What it costs in practice
Typical request (1,200 in + 400 out tokens)$0.0003
1,000 requests / month$0.3224/mo10,000 requests / month$3.22/mo100,000 requests / month$32.24/moEstimate your own workloadOpens the calculator with Qwen3 VL 8B Instruct preloaded — adjust volume and token counts there.
Price history
List price per 1M tokens since we started tracking in Jun 2026. Each step marks a price change — hover a point for the exact date and rate.
What it excels at
Handles multimodal inputs for tasks such as visual question answering and document analysis at low inference cost.
The business tradeoff
8B scale restricts performance on complex multi-step reasoning compared with larger models in the same family; open-weights deployment introduces variable latency and throughput.
Head-to-head
Vendor list rates, as of Jul 28, 2026 · source: openrouter · per-request examples assume 1,200 input + 400 output tokens.