Qwen3.6 35B A3B pricing & benchmarks
Qwen3.6 35B A3B is available from Alibaba at $0.140 per million input tokens and $1.00 per million output tokens ($0.355 blended at 3:1). It accepts up to 262,144 tokens of context. Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
Pricing
Per million tokens ยท OpenRouter
Input$0.140
Output$1.00
Blended (3:1)$0.355
Specification
qwen/qwen3.6-35b-a3b
Context window262K
Max output262K
ReleasedApr 2026
LicenceOpen weights
Modalitiesin:text, in:image, in:video, out:text
No independently-run benchmark scores are published for Qwen3.6 35B A3B yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen3.6 35B A3B
9 hosts ยท 2.2ร between cheapest and dearest. Blended at 3:1. โ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
| Host | Blended $/1M | In / Out | Context | Weights | Uptime 24h |
|---|---|---|---|---|---|
| Venice | $0.311 | $0.098 / $0.950 | 256K | fp8 | 99.9% |
| DeepInfra | $0.313 | $0.100 / $0.950 | 262K | fp8 | 99.1% |
| AkashML | $0.355 | $0.140 / $1.00 | 262K | fp8 | 98.7% |
| Parasail | $0.362 | $0.150 / $1.00 | 262K | fp8 | 96.0% |
| AtlasCloud | $0.418 | $0.186 / $1.11 | 262K | fp8 | 99.6% |
| Phala | $0.468 | $0.200 / $1.27 | 262K | โ | 95.2% |
| CoreWeave | $0.500 | $0.250 / $1.25 | 262K | fp8 | 99.1% |
| SiliconFlow | $0.550 | $0.200 / $1.60 | 262K | fp8 | 98.2% |
| Io Net | $0.690 | $0.290 / $1.89 | 262K | fp8 | 98.2% |
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 โ the score above is not host-specific.
Compare Qwen3.6 35B A3B
FAQ
How much does Qwen3.6 35B A3B cost?
Qwen3.6 35B A3B costs $0.140 per million input tokens and $1.00 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.355 per million tokens.
What is Qwen3.6 35B A3B's context window?
Qwen3.6 35B A3B accepts up to 262,144 tokens of context, and can generate up to 262,144 output tokens.