R1 Distill Llama 70B pricing & benchmarks
R1 Distill Llama 70B is available from DeepSeek at $0.800 per million input tokens and $0.800 per million output tokens ($0.800 blended at 3:1). It accepts up to 8,192 tokens of context. DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation technβ¦
Pricing
Per million tokens Β· OpenRouter
Input$0.800
Output$0.800
Blended (3:1)$0.800
Specification
deepseek/deepseek-r1-distill-llama-70b
Context window8K
Max output8K
ReleasedJan 2025
LicenceOpen weights
Modalitiesin:text, out:text
Benchmarks
Independently run β not vendor self-reports. Every score links its source.
| Benchmark | Score | Config | Run | Source |
|---|---|---|---|---|
| GPQA Diamond | 55.7% Β±3.0 | β | 2025-03-10 | Independently run Β· Epoch AI |
| OTIS Mock AIME | 51.4% Β±6.1 | β | 2025-03-07 | Independently run Β· Epoch AI |
Scores published by Epoch AI, 'AI Benchmarking Hub', used under CC BY 4.0.
Cheaper models that score at least as well on GPQA Diamond
Strictly lower blended price AND an equal-or-higher GPQA Diamond score. One benchmark is one dimension β a model that wins here may still be weaker at your specific task.
| Model | Provider | Blended $/1M | GPQA Diamond | |
|---|---|---|---|---|
| gpt-oss-120b | OpenAI | $0.070 (β91%) | 75.8% | Compare β |
| Phi 4 | Microsoft | $0.088 (β89%) | 56.1% | Compare β |
| DeepSeek V4 Flash 0731 | DeepSeek | $0.113 (β86%) | 91.0% | Compare β |
| GPT-5 Nano | OpenAI | $0.138 (β83%) | 69.4% | Compare β |
| Gemma 4 31B | $0.160 (β80%) | 75.8% | Compare β |
Compare R1 Distill Llama 70B
FAQ
How much does R1 Distill Llama 70B cost?
R1 Distill Llama 70B costs $0.800 per million input tokens and $0.800 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.800 per million tokens.
What is R1 Distill Llama 70B's context window?
R1 Distill Llama 70B accepts up to 8,192 tokens of context, and can generate up to 8,192 output tokens.
How good is R1 Distill Llama 70B on benchmarks?
R1 Distill Llama 70B scores 55.7% on GPQA Diamond, independently run and published by Epoch AI. 2 benchmark results are listed on this page, each with its run date and source.
Is there a cheaper model as good as R1 Distill Llama 70B?
Yes β 5 models in our catalogue cost less per token than R1 Distill Llama 70B and score at least as high on the same benchmark. The cheapest is gpt-oss-120b at $0.070 per million tokens blended.