token.appAlibaba › Qwen3 VL 8B Instruct

Qwen3 VL 8B Instruct pricing & benchmarks

Qwen3 VL 8B Instruct is available from Alibaba at $0.117 per million input tokens and $0.455 per million output tokens ($0.202 blended at 3:1). It accepts up to 262,144 tokens of context. Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-…

Pricing
Per million tokens · OpenRouter
Input$0.117
Output$0.455
Blended (3:1)$0.202
Specification
qwen/qwen3-vl-8b-instruct
Context window262K
Max output33K
ReleasedOct 2025
LicenceOpen weights
Modalitiesin:text, in:image, out:text
No independently-run benchmark scores are published for Qwen3 VL 8B Instruct yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen3 VL 8B Instruct
2 hosts · 1.9× between cheapest and dearest. Blended at 3:1. ⚠ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
HostBlended $/1MIn / OutContextWeightsUptime 24h
Alibaba
⚠ 131K context, not 262K
$0.202 $0.117 / $0.455 131K fp8 100.0%
Parasail $0.375 $0.250 / $0.750 262K bf16 99.2%
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 — the score above is not host-specific.
Compare Qwen3 VL 8B Instruct
FAQ
How much does Qwen3 VL 8B Instruct cost?
Qwen3 VL 8B Instruct costs $0.117 per million input tokens and $0.455 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.202 per million tokens.
What is Qwen3 VL 8B Instruct's context window?
Qwen3 VL 8B Instruct accepts up to 262,144 tokens of context, and can generate up to 32,768 output tokens.