Qwen2.5 7B Instruct pricing & benchmarks
Qwen2.5 7B Instruct is available from Alibaba at $0.100 per million input tokens and $0.200 per million output tokens ($0.125 blended at 3:1). It accepts up to 32,768 tokens of context. Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...
Pricing
Per million tokens ยท OpenRouter
Input$0.100
Output$0.200
Blended (3:1)$0.125
Specification
qwen/qwen-2.5-7b-instruct
Context window33K
Max output33K
ReleasedOct 2024
LicenceOpen weights
Modalitiesin:text, out:text
No independently-run benchmark scores are published for Qwen2.5 7B Instruct yet. We show nothing rather than reprinting an unverified figure.
Where to run Qwen2.5 7B Instruct
2 hosts ยท 2.4ร between cheapest and dearest. Blended at 3:1. โ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
| Host | Blended $/1M | In / Out | Context | Weights | Uptime 24h |
|---|---|---|---|---|---|
| Phala | $0.125 | $0.100 / $0.200 | 33K | โ | 99.6% |
| Together | $0.300 | $0.300 / $0.300 | 33K | fp8 | 99.9% |
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 โ the score above is not host-specific.
Compare Qwen2.5 7B Instruct
FAQ
How much does Qwen2.5 7B Instruct cost?
Qwen2.5 7B Instruct costs $0.100 per million input tokens and $0.200 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.125 per million tokens.
What is Qwen2.5 7B Instruct's context window?
Qwen2.5 7B Instruct accepts up to 32,768 tokens of context, and can generate up to 32,768 output tokens.