Qwen3.5-9B pricing & benchmarks
Qwen3.5-9B is available from Alibaba at $0.100 per million input tokens and $0.150 per million output tokens ($0.113 blended at 3:1). It accepts up to 262,144 tokens of context. Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
Pricing
Per million tokens ยท OpenRouter
Input$0.100
Output$0.150
Blended (3:1)$0.113
Specification
qwen/qwen3.5-9b
Context window262K
Max output262K
ReleasedMar 2026
LicenceOpen weights
Modalitiesin:text, in:image, in:video, out:text
No independently-run benchmark scores are published for Qwen3.5-9B yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen3.5-9B
5 hosts ยท 1.7ร between cheapest and dearest. Blended at 3:1. โ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
| Host | Blended $/1M | In / Out | Context | Weights | Uptime 24h |
|---|---|---|---|---|---|
| SiliconFlow | $0.112 | $0.100 / $0.150 | 262K | fp8 | 99.1% |
| DeepInfra | $0.112 | $0.100 / $0.150 | 262K | bf16 | 98.2% |
| Venice | $0.112 | $0.100 / $0.150 | 256K | fp8 | 96.7% |
| Parasail | $0.138 | $0.100 / $0.250 | 262K | bf16 | 97.1% |
| Together | $0.190 | $0.170 / $0.250 | 262K | โ | 99.4% |
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 โ the score above is not host-specific.
Compare Qwen3.5-9B
FAQ
How much does Qwen3.5-9B cost?
Qwen3.5-9B costs $0.100 per million input tokens and $0.150 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.113 per million tokens.
What is Qwen3.5-9B's context window?
Qwen3.5-9B accepts up to 262,144 tokens of context, and can generate up to 262,144 output tokens.