token.app โ€บ Alibaba โ€บ Qwen3 32B

Qwen3 32B pricing & benchmarks

Qwen3 32B is available from Alibaba at $0.080 per million input tokens and $0.280 per million output tokens ($0.130 blended at 3:1). It accepts up to 131,072 tokens of context. Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

Pricing
Per million tokens ยท OpenRouter
Input$0.080
Output$0.280
Blended (3:1)$0.130
Specification
qwen/qwen3-32b
Context window131K
Max output16K
ReleasedApr 2025
LicenceOpen weights
Modalitiesin:text, out:text
No independently-run benchmark scores are published for Qwen3 32B yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen3 32B
5 hosts ยท 2.8ร— between cheapest and dearest. Blended at 3:1. โš  marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
HostBlended $/1MIn / OutContextWeightsUptime 24h
DeepInfra
โš  41K context, not 131K
$0.130 $0.080 / $0.280 41K fp8 99.9%
Nebius (base)
โš  41K context, not 131K
$0.150 $0.100 / $0.300 41K fp8 98.7%
Alibaba $0.182 $0.104 / $0.416 131K fp8 0.0%
SiliconFlow $0.248 $0.140 / $0.570 131K fp8 99.1%
Groq $0.365 $0.290 / $0.590 131K โ€” 100.0%
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 โ€” the score above is not host-specific.
Compare Qwen3 32B
FAQ
How much does Qwen3 32B cost?
Qwen3 32B costs $0.080 per million input tokens and $0.280 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.130 per million tokens.
What is Qwen3 32B's context window?
Qwen3 32B accepts up to 131,072 tokens of context, and can generate up to 16,384 output tokens.