token.app โ€บ Alibaba โ€บ Qwen3 235B A22B Thinking 2507

Qwen3 235B A22B Thinking 2507 pricing & benchmarks

Qwen3 235B A22B Thinking 2507 is available from Alibaba at $0.230 per million input tokens and $2.30 per million output tokens ($0.747 blended at 3:1). It accepts up to 262,144 tokens of context. Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...

Pricing
Per million tokens ยท OpenRouter
Input$0.230
Output$2.30
Blended (3:1)$0.747
Specification
qwen/qwen3-235b-a22b-thinking-2507
Context window262K
Max outputโ€”
ReleasedJul 2025
LicenceOpen weights
Modalitiesin:text, out:text
No independently-run benchmark scores are published for Qwen3 235B A22B Thinking 2507 yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen3 235B A22B Thinking 2507
4 hosts ยท 1.6ร— between cheapest and dearest. Blended at 3:1. โš  marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
HostBlended $/1MIn / OutContextWeightsUptime 24h
DeepInfra $0.747 $0.230 / $2.30 262K fp8 97.9%
Alibaba
โš  131K context, not 262K
$0.747 $0.230 / $2.30 131K fp8 98.8%
Novita
โš  131K context, not 262K
$0.975 $0.300 / $3.00 131K fp8 88.0%
Venice
โš  128K context, not 262K
$1.21 $0.450 / $3.50 128K fp8 89.6%
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 โ€” the score above is not host-specific.
Compare Qwen3 235B A22B Thinking 2507
FAQ
How much does Qwen3 235B A22B Thinking 2507 cost?
Qwen3 235B A22B Thinking 2507 costs $0.230 per million input tokens and $2.30 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.747 per million tokens.
What is Qwen3 235B A22B Thinking 2507's context window?
Qwen3 235B A22B Thinking 2507 accepts up to 262,144 tokens of context.