token.appAlibaba › Qwen3 Next 80B A3B Instruct

Qwen3 Next 80B A3B Instruct pricing & benchmarks

Qwen3 Next 80B A3B Instruct is available from Alibaba at $0.090 per million input tokens and $1.10 per million output tokens ($0.343 blended at 3:1). It accepts up to 262,144 tokens of context. Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingu…

Pricing
Per million tokens · OpenRouter
Input$0.090
Output$1.10
Blended (3:1)$0.343
Specification
qwen/qwen3-next-80b-a3b-instruct
Context window262K
Max output16K
ReleasedSep 2025
LicenceOpen weights
Modalitiesin:text, out:text
No independently-run benchmark scores are published for Qwen3 Next 80B A3B Instruct yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen3 Next 80B A3B Instruct
5 hosts · 1.8× between cheapest and dearest. Blended at 3:1. ⚠ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
HostBlended $/1MIn / OutContextWeightsUptime 24h
Alibaba
⚠ 131K context, not 262K
$0.268 $0.098 / $0.780 131K fp8 100.0%
DeepInfra $0.343 $0.090 / $1.10 262K fp8 99.4%
Parasail $0.350 $0.100 / $1.10 262K fp8 97.4%
Google (global) $0.412 $0.150 / $1.20 262K 100.0%
Novita
⚠ 131K context, not 262K
$0.487 $0.150 / $1.50 131K bf16 90.9%
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 — the score above is not host-specific.
Compare Qwen3 Next 80B A3B Instruct
FAQ
How much does Qwen3 Next 80B A3B Instruct cost?
Qwen3 Next 80B A3B Instruct costs $0.090 per million input tokens and $1.10 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.343 per million tokens.
What is Qwen3 Next 80B A3B Instruct's context window?
Qwen3 Next 80B A3B Instruct accepts up to 262,144 tokens of context, and can generate up to 16,384 output tokens.