token.app โ€บ Alibaba โ€บ Qwen3.6 35B A3B

Qwen3.6 35B A3B pricing & benchmarks

Qwen3.6 35B A3B is available from Alibaba at $0.140 per million input tokens and $1.00 per million output tokens ($0.355 blended at 3:1). It accepts up to 262,144 tokens of context. Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...

Pricing
Per million tokens ยท OpenRouter
Input$0.140
Output$1.00
Blended (3:1)$0.355
Specification
qwen/qwen3.6-35b-a3b
Context window262K
Max output262K
ReleasedApr 2026
LicenceOpen weights
Modalitiesin:text, in:image, in:video, out:text
No independently-run benchmark scores are published for Qwen3.6 35B A3B yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen3.6 35B A3B
9 hosts ยท 2.2ร— between cheapest and dearest. Blended at 3:1. โš  marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
HostBlended $/1MIn / OutContextWeightsUptime 24h
Venice $0.311 $0.098 / $0.950 256K fp8 99.9%
DeepInfra $0.313 $0.100 / $0.950 262K fp8 99.1%
AkashML $0.355 $0.140 / $1.00 262K fp8 98.7%
Parasail $0.362 $0.150 / $1.00 262K fp8 96.0%
AtlasCloud $0.418 $0.186 / $1.11 262K fp8 99.6%
Phala $0.468 $0.200 / $1.27 262K โ€” 95.2%
CoreWeave $0.500 $0.250 / $1.25 262K fp8 99.1%
SiliconFlow $0.550 $0.200 / $1.60 262K fp8 98.2%
Io Net $0.690 $0.290 / $1.89 262K fp8 98.2%
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 โ€” the score above is not host-specific.
Compare Qwen3.6 35B A3B
FAQ
How much does Qwen3.6 35B A3B cost?
Qwen3.6 35B A3B costs $0.140 per million input tokens and $1.00 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.355 per million tokens.
What is Qwen3.6 35B A3B's context window?
Qwen3.6 35B A3B accepts up to 262,144 tokens of context, and can generate up to 262,144 output tokens.