token.app โ€บ Alibaba โ€บ Qwen3.5-35B-A3B

Qwen3.5-35B-A3B pricing & benchmarks

Qwen3.5-35B-A3B is available from Alibaba at $0.140 per million input tokens and $1.00 per million output tokens ($0.355 blended at 3:1). It accepts up to 262,144 tokens of context. The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...

Pricing
Per million tokens ยท OpenRouter
Input$0.140
Output$1.00
Blended (3:1)$0.355
Specification
qwen/qwen3.5-35b-a3b
Context window262K
Max output262K
ReleasedFeb 2026
LicenceOpen weights
Modalitiesin:text, in:image, in:video, out:text
No independently-run benchmark scores are published for Qwen3.5-35B-A3B yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen3.5-35B-A3B
9 hosts ยท 1.8ร— between cheapest and dearest. Blended at 3:1. โš  marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
HostBlended $/1MIn / OutContextWeightsUptime 24h
DeepInfra $0.355 $0.140 / $1.00 262K fp8 99.5%
AkashML $0.355 $0.140 / $1.00 262K fp8 99.8%
Parasail $0.362 $0.150 / $1.00 262K fp8 84.3%
Alibaba $0.447 $0.163 / $1.30 262K fp8 100.0%
CoreWeave $0.500 $0.250 / $1.25 262K fp8 100.0%
NextBit $0.515 $0.220 / $1.40 262K fp8 0.0%
Venice $0.547 $0.313 / $1.25 256K โ€” 95.9%
AtlasCloud $0.619 $0.225 / $1.80 262K fp8 99.8%
SiliconFlow $0.630 $0.240 / $1.80 262K fp8 95.8%
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 โ€” the score above is not host-specific.
Compare Qwen3.5-35B-A3B
FAQ
How much does Qwen3.5-35B-A3B cost?
Qwen3.5-35B-A3B costs $0.140 per million input tokens and $1.00 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.355 per million tokens.
What is Qwen3.5-35B-A3B's context window?
Qwen3.5-35B-A3B accepts up to 262,144 tokens of context, and can generate up to 262,144 output tokens.