Qwen3 VL 8B Instruct pricing & benchmarks
Qwen3 VL 8B Instruct is available from Alibaba at $0.117 per million input tokens and $0.455 per million output tokens ($0.202 blended at 3:1). It accepts up to 262,144 tokens of context. Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-…
Pricing
Per million tokens · OpenRouter
Input$0.117
Output$0.455
Blended (3:1)$0.202
Specification
qwen/qwen3-vl-8b-instruct
Context window262K
Max output33K
ReleasedOct 2025
LicenceOpen weights
Modalitiesin:text, in:image, out:text
No independently-run benchmark scores are published for Qwen3 VL 8B Instruct yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen3 VL 8B Instruct
2 hosts · 1.9× between cheapest and dearest. Blended at 3:1. ⚠ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
| Host | Blended $/1M | In / Out | Context | Weights | Uptime 24h |
|---|---|---|---|---|---|
| Alibaba ⚠ 131K context, not 262K |
$0.202 | $0.117 / $0.455 | 131K | fp8 | 100.0% |
| Parasail | $0.375 | $0.250 / $0.750 | 262K | bf16 | 99.2% |
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 — the score above is not host-specific.
Compare Qwen3 VL 8B Instruct
Qwen3 VL 8B Instruct vs DeepSeek V4 Flash 0731
Qwen3 VL 8B Instruct vs Claude Opus 5
Qwen3 VL 8B Instruct vs Gemini 3.6 Flash
Qwen3 VL 8B Instruct vs Kimi K3
Qwen3 VL 8B Instruct vs GPT-5.6 Luna
Qwen3 VL 8B Instruct vs GPT-5.6 Terra
Qwen3 VL 8B Instruct vs GPT-5.6 Sol
Qwen3 VL 8B Instruct vs Grok 4.5
FAQ
How much does Qwen3 VL 8B Instruct cost?
Qwen3 VL 8B Instruct costs $0.117 per million input tokens and $0.455 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.202 per million tokens.
What is Qwen3 VL 8B Instruct's context window?
Qwen3 VL 8B Instruct accepts up to 262,144 tokens of context, and can generate up to 32,768 output tokens.