Qwen3 30B A3B pricing & benchmarks
Qwen3 30B A3B is available from Alibaba at $0.120 per million input tokens and $0.500 per million output tokens ($0.215 blended at 3:1). It accepts up to 131,072 tokens of context. Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...
Pricing
Per million tokens · OpenRouter
Input$0.120
Output$0.500
Blended (3:1)$0.215
Specification
qwen/qwen3-30b-a3b
Context window131K
Max output16K
ReleasedApr 2025
LicenceOpen weights
Modalitiesin:text, out:text
No independently-run benchmark scores are published for Qwen3 30B A3B yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen3 30B A3B
2 hosts · 1.1× between cheapest and dearest. Blended at 3:1. ⚠ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
| Host | Blended $/1M | In / Out | Context | Weights | Uptime 24h |
|---|---|---|---|---|---|
| DeepInfra ⚠ 41K context, not 131K |
$0.215 | $0.120 / $0.500 | 41K | fp8 | 99.9% |
| Alibaba | $0.228 | $0.130 / $0.520 | 131K | fp8 | 99.9% |
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 — the score above is not host-specific.
Compare Qwen3 30B A3B
FAQ
How much does Qwen3 30B A3B cost?
Qwen3 30B A3B costs $0.120 per million input tokens and $0.500 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.215 per million tokens.
What is Qwen3 30B A3B's context window?
Qwen3 30B A3B accepts up to 131,072 tokens of context, and can generate up to 16,384 output tokens.