token.app › Compare › Llama 3.3 70B Instruct vs gpt-oss-120b
Llama 3.3 70B Instruct vs gpt-oss-120b
gpt-oss-120b leads on 2 of 2 shared benchmarks. gpt-oss-120b is 2.2× cheaper per token on a blended 3:1 basis.
Head to head
Blended price uses a 3:1 input:output mix
| Llama 3.3 70B Instruct | gpt-oss-120b | |
|---|---|---|
| Provider | Meta | OpenAI |
| Input $/1M | $0.100 | $0.037 |
| Output $/1M | $0.320 | $0.170 |
| Blended $/1M | $0.155 | $0.070 |
| Context window | 131K | 131K |
| Max output | 16K | 131K |
| Released | Dec 2024 | Aug 2025 |
| Licence | Open weights | Open weights |
| GPQA Diamond | 47.4% | 75.8% |
| OTIS Mock AIME | 5.1% | 88.9% |
Benchmark scores independently run and published by Epoch AI, 'AI Benchmarking Hub', used under CC BY 4.0. Only benchmarks with a score for BOTH models are compared.
Choose Llama 3.3 70B Instruct if…
- No clear advantage on the data we hold.
Choose gpt-oss-120b if…
- Costs less per token — $0.070 vs $0.155 blended
- Higher GPQA Diamond (75.8% vs 47.4%)
- Higher OTIS Mock AIME (88.9% vs 5.1%)
- Newer model (released Aug 2025)
FAQ
Which is better, Llama 3.3 70B Instruct or gpt-oss-120b?
Across the 2 benchmarks both models have independently-run scores for, Llama 3.3 70B Instruct leads on 0 and gpt-oss-120b leads on 2. Benchmarks are one input — price and context window are on this page too.
Is Llama 3.3 70B Instruct cheaper than gpt-oss-120b?
gpt-oss-120b is cheaper: $0.070 versus $0.155 per million tokens blended at 3:1 input:output. Input alone: $0.100 vs $0.037. Output alone: $0.320 vs $0.170.
What context windows do Llama 3.3 70B Instruct and gpt-oss-120b have?
Llama 3.3 70B Instruct: 131,072 tokens. gpt-oss-120b: 131,072 tokens.
Who makes Llama 3.3 70B Instruct and gpt-oss-120b?
Llama 3.3 70B Instruct is made by Meta. gpt-oss-120b is made by OpenAI.
Related comparisons
Llama 3.3 70B Instruct full details
gpt-oss-120b full details
Llama 3.3 70B Instruct vs Ling 3.0 Tiny (free)
Llama 3.3 70B Instruct vs Muse Spark 1.2
Llama 3.3 70B Instruct vs Qwen3.8 Max
Llama 3.3 70B Instruct vs DeepSeek V4 Flash Latest
Llama 3.3 70B Instruct vs DeepSeek V4 Flash 0731
Llama 3.3 70B Instruct vs Inkling Small