token.app โ€บ Alibaba โ€บ Qwen3.5 397B A17B

Qwen3.5 397B A17B pricing & benchmarks

Qwen3.5 397B A17B is available from Alibaba at $0.390 per million input tokens and $2.34 per million output tokens ($0.877 blended at 3:1). It accepts up to 262,144 tokens of context. The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...

Pricing
Per million tokens ยท OpenRouter
Input$0.390
Output$2.34
Blended (3:1)$0.877
Specification
qwen/qwen3.5-397b-a17b
Context window262K
Max output66K
ReleasedFeb 2026
LicenceOpen weights
Modalitiesin:text, in:image, in:video, out:text
No independently-run benchmark scores are published for Qwen3.5 397B A17B yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen3.5 397B A17B
11 hosts ยท 2.4ร— between cheapest and dearest. Blended at 3:1. โš  marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
HostBlended $/1MIn / OutContextWeightsUptime 24h
DigitalOcean
โš  131K context, not 262K
$0.708 $0.302 / $1.93 131K โ€” 98.0%
Alibaba $0.877 $0.390 / $2.34 262K fp8 99.9%
Chutes $1.09 $0.450 / $3.00 262K fp8 98.7%
DeepInfra $1.09 $0.450 / $3.00 262K fp8 96.9%
Parasail $1.27 $0.500 / $3.60 262K fp8 98.2%
AtlasCloud $1.29 $0.550 / $3.50 262K fp8 96.4%
Phala $1.29 $0.550 / $3.50 262K โ€” 97.5%
Novita $1.35 $0.600 / $3.60 262K โ€” 99.6%
StreamLake $1.35 $0.600 / $3.60 256K โ€” 98.6%
GMICloud $1.35 $0.600 / $3.60 262K fp8 0.0%
Venice
โš  128K context, not 262K
$1.69 $0.750 / $4.50 128K โ€” 96.8%
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 โ€” the score above is not host-specific.
Compare Qwen3.5 397B A17B
FAQ
How much does Qwen3.5 397B A17B cost?
Qwen3.5 397B A17B costs $0.390 per million input tokens and $2.34 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.877 per million tokens.
What is Qwen3.5 397B A17B's context window?
Qwen3.5 397B A17B accepts up to 262,144 tokens of context, and can generate up to 65,536 output tokens.