Qwen3 VL 30B A3B Instruct pricing & benchmarks
Qwen3 VL 30B A3B Instruct is available from Alibaba at $0.150 per million input tokens and $0.600 per million output tokens ($0.262 blended at 3:1). It accepts up to 262,144 tokens of context. Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception.…
Pricing
Per million tokens · OpenRouter
Input$0.150
Output$0.600
Blended (3:1)$0.262
Specification
qwen/qwen3-vl-30b-a3b-instruct
Context window262K
Max output16K
ReleasedOct 2025
LicenceOpen weights
Modalitiesin:text, in:image, out:text
No independently-run benchmark scores are published for Qwen3 VL 30B A3B Instruct yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen3 VL 30B A3B Instruct
4 hosts · 2.1× between cheapest and dearest. Blended at 3:1. ⚠ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
| Host | Blended $/1M | In / Out | Context | Weights | Uptime 24h |
|---|---|---|---|---|---|
| Alibaba ⚠ 131K context, not 262K |
$0.228 | $0.130 / $0.520 | 131K | fp8 | 99.8% |
| DeepInfra | $0.262 | $0.150 / $0.600 | 262K | fp8 | 99.4% |
| Novita ⚠ 131K context, not 262K |
$0.325 | $0.200 / $0.700 | 131K | bf16 | 97.3% |
| SiliconFlow | $0.467 | $0.290 / $1.00 | 262K | fp8 | 96.2% |
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 — the score above is not host-specific.
Compare Qwen3 VL 30B A3B Instruct
Qwen3 VL 30B A3B Instruct vs DeepSeek V4 Flash 0731
Qwen3 VL 30B A3B Instruct vs Claude Opus 5
Qwen3 VL 30B A3B Instruct vs Gemini 3.6 Flash
Qwen3 VL 30B A3B Instruct vs Kimi K3
Qwen3 VL 30B A3B Instruct vs GPT-5.6 Luna
Qwen3 VL 30B A3B Instruct vs GPT-5.6 Terra
Qwen3 VL 30B A3B Instruct vs GPT-5.6 Sol
Qwen3 VL 30B A3B Instruct vs Grok 4.5
FAQ
How much does Qwen3 VL 30B A3B Instruct cost?
Qwen3 VL 30B A3B Instruct costs $0.150 per million input tokens and $0.600 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.262 per million tokens.
What is Qwen3 VL 30B A3B Instruct's context window?
Qwen3 VL 30B A3B Instruct accepts up to 262,144 tokens of context, and can generate up to 16,384 output tokens.