DeepSeek V3.1 pricing & benchmarks
DeepSeek V3.1 is available from DeepSeek at $0.250 per million input tokens and $0.950 per million output tokens ($0.425 blended at 3:1). It accepts up to 163,840 tokens of context. DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context...
Pricing
Per million tokens ยท OpenRouter
Input$0.250
Output$0.950
Blended (3:1)$0.425
Specification
deepseek/deepseek-chat-v3.1
Context window164K
Max output33K
ReleasedAug 2025
LicenceOpen weights
Modalitiesin:text, out:text
No independently-run benchmark scores are published for DeepSeek V3.1 yet. We show nothing rather than reprinting an unverified figure.
Where to run DeepSeek V3.1
8 hosts ยท 2.1ร between cheapest and dearest. Blended at 3:1. โ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
| Host | Blended $/1M | In / Out | Context | Weights | Uptime 24h |
|---|---|---|---|---|---|
| DeepInfra | $0.425 | $0.250 / $0.950 | 164K | fp4 | 99.9% |
| Novita โ 131K context, not 164K |
$0.453 | $0.270 / $1.00 | 131K | fp8 | 100.0% |
| SiliconFlow | $0.453 | $0.270 / $1.00 | 164K | fp8 | 97.5% |
| AtlasCloud โ 131K context, not 164K |
$0.462 | $0.300 / $0.950 | 131K | fp8 | 99.9% |
| CoreWeave | $0.825 | $0.550 / $1.65 | 161K | fp8 | 100.0% |
| SambaNova โ 131K context, not 164K |
$0.863 | $0.650 / $1.50 | 131K | fp8 | 99.5% |
| Google (us-west2) | $0.875 | $0.600 / $1.70 | 164K | โ | 97.4% |
| Mara โ 131K context, not 164K |
$0.875 | $0.600 / $1.70 | 131K | โ | 76.9% |
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 โ the score above is not host-specific.
Compare DeepSeek V3.1
FAQ
How much does DeepSeek V3.1 cost?
DeepSeek V3.1 costs $0.250 per million input tokens and $0.950 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.425 per million tokens.
What is DeepSeek V3.1's context window?
DeepSeek V3.1 accepts up to 163,840 tokens of context, and can generate up to 32,768 output tokens.