token.appAlibaba › Qwen3 VL 235B A22B Thinking

Qwen3 VL 235B A22B Thinking pricing & benchmarks

Qwen3 VL 235B A22B Thinking is available from Alibaba at $0.400 per million input tokens and $4.00 per million output tokens ($1.30 blended at 3:1). It accepts up to 131,072 tokens of context. Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STEM and math....

Pricing
Per million tokens · OpenRouter
Input$0.400
Output$4.00
Blended (3:1)$1.30
Specification
qwen/qwen3-vl-235b-a22b-thinking
Context window131K
Max output33K
ReleasedSep 2025
LicenceOpen weights
Modalitiesin:text, in:image, out:text
No independently-run benchmark scores are published for Qwen3 VL 235B A22B Thinking yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen3 VL 235B A22B Thinking
2 hosts · 1.3× between cheapest and dearest. Blended at 3:1. ⚠ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
HostBlended $/1MIn / OutContextWeightsUptime 24h
Alibaba $1.30 $0.400 / $4.00 131K fp8 91.1%
Novita $1.72 $0.980 / $3.95 131K bf16 99.9%
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 — the score above is not host-specific.
Compare Qwen3 VL 235B A22B Thinking
FAQ
How much does Qwen3 VL 235B A22B Thinking cost?
Qwen3 VL 235B A22B Thinking costs $0.400 per million input tokens and $4.00 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $1.30 per million tokens.
What is Qwen3 VL 235B A22B Thinking's context window?
Qwen3 VL 235B A22B Thinking accepts up to 131,072 tokens of context, and can generate up to 32,768 output tokens.