token.app › Compare › R1 Distill Llama 70B vs gpt-oss-120b
R1 Distill Llama 70B vs gpt-oss-120b
gpt-oss-120b leads on 2 of 2 shared benchmarks. gpt-oss-120b is 11× cheaper per token on a blended 3:1 basis. gpt-oss-120b accepts the larger context window (131K).
Head to head
Blended price uses a 3:1 input:output mix
| R1 Distill Llama 70B | gpt-oss-120b | |
|---|---|---|
| Provider | DeepSeek | OpenAI |
| Input $/1M | $0.800 | $0.037 |
| Output $/1M | $0.800 | $0.170 |
| Blended $/1M | $0.800 | $0.070 |
| Context window | 8K | 131K |
| Max output | 8K | 131K |
| Released | Jan 2025 | Aug 2025 |
| Licence | Open weights | Open weights |
| GPQA Diamond | 55.7% | 75.8% |
| OTIS Mock AIME | 51.4% | 88.9% |
Benchmark scores independently run and published by Epoch AI, 'AI Benchmarking Hub', used under CC BY 4.0. Only benchmarks with a score for BOTH models are compared.
Choose R1 Distill Llama 70B if…
- No clear advantage on the data we hold.
Choose gpt-oss-120b if…
- Costs less per token — $0.070 vs $0.800 blended
- Higher GPQA Diamond (75.8% vs 55.7%)
- Higher OTIS Mock AIME (88.9% vs 51.4%)
- Larger context window (131K)
- Newer model (released Aug 2025)
FAQ
Which is better, R1 Distill Llama 70B or gpt-oss-120b?
Across the 2 benchmarks both models have independently-run scores for, R1 Distill Llama 70B leads on 0 and gpt-oss-120b leads on 2. Benchmarks are one input — price and context window are on this page too.
Is R1 Distill Llama 70B cheaper than gpt-oss-120b?
gpt-oss-120b is cheaper: $0.070 versus $0.800 per million tokens blended at 3:1 input:output. Input alone: $0.800 vs $0.037. Output alone: $0.800 vs $0.170.
What context windows do R1 Distill Llama 70B and gpt-oss-120b have?
R1 Distill Llama 70B: 8,192 tokens. gpt-oss-120b: 131,072 tokens.
Who makes R1 Distill Llama 70B and gpt-oss-120b?
R1 Distill Llama 70B is made by DeepSeek. gpt-oss-120b is made by OpenAI.
Related comparisons
R1 Distill Llama 70B full details
gpt-oss-120b full details
R1 Distill Llama 70B vs Ling 3.0 Tiny (free)
R1 Distill Llama 70B vs Muse Spark 1.2
R1 Distill Llama 70B vs Qwen3.8 Max
R1 Distill Llama 70B vs DeepSeek V4 Flash Latest
R1 Distill Llama 70B vs DeepSeek V4 Flash 0731
R1 Distill Llama 70B vs Inkling Small