Qwen3 VL 32B Instruct pricing & benchmarks
Qwen3 VL 32B Instruct is available from Alibaba at $0.104 per million input tokens and $0.416 per million output tokens ($0.182 blended at 3:1). It accepts up to 131,072 tokens of context. Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced tex…
Pricing
Per million tokens · OpenRouter
Input$0.104
Output$0.416
Blended (3:1)$0.182
Specification
qwen/qwen3-vl-32b-instruct
Context window131K
Max output33K
ReleasedOct 2025
LicenceOpen weights
Modalitiesin:text, in:image, out:text
No independently-run benchmark scores are published for Qwen3 VL 32B Instruct yet. We show nothing rather than reprinting an unverified figure.
Compare Qwen3 VL 32B Instruct
Qwen3 VL 32B Instruct vs DeepSeek V4 Flash 0731
Qwen3 VL 32B Instruct vs Claude Opus 5
Qwen3 VL 32B Instruct vs Gemini 3.6 Flash
Qwen3 VL 32B Instruct vs Kimi K3
Qwen3 VL 32B Instruct vs GPT-5.6 Luna
Qwen3 VL 32B Instruct vs GPT-5.6 Terra
Qwen3 VL 32B Instruct vs GPT-5.6 Sol
Qwen3 VL 32B Instruct vs Grok 4.5
FAQ
How much does Qwen3 VL 32B Instruct cost?
Qwen3 VL 32B Instruct costs $0.104 per million input tokens and $0.416 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.182 per million tokens.
What is Qwen3 VL 32B Instruct's context window?
Qwen3 VL 32B Instruct accepts up to 131,072 tokens of context, and can generate up to 32,768 output tokens.