token.app โ€บ OpenAI โ€บ gpt-oss-20b

gpt-oss-20b pricing & benchmarks

gpt-oss-20b is available from OpenAI at $0.030 per million input tokens and $0.130 per million output tokens ($0.055 blended at 3:1). It accepts up to 131,072 tokens of context. gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

Pricing
Per million tokens ยท OpenRouter
Input$0.030
Output$0.130
Blended (3:1)$0.055
Specification
openai/gpt-oss-20b
Context window131K
Max output131K
ReleasedAug 2025
LicenceOpen weights
Modalitiesin:text, out:text
No independently-run benchmark scores are published for gpt-oss-20b yet. We show nothing rather than reprinting an unverified figure.
Where to run gpt-oss-20b
12 options across 11 providers ยท 2.4ร— between cheapest and dearest. Blended at 3:1. โš  marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
HostBlended $/1MIn / OutContextWeightsUptime 24h
CoreWeave $0.055 $0.030 / $0.130 131K fp4 99.2%
DeepInfra $0.058 $0.030 / $0.140 131K bf16 99.8%
Parasail $0.060 $0.030 / $0.150 131K fp4 81.5%
Novita $0.068 $0.040 / $0.150 131K fp4 95.9%
Phala $0.068 $0.040 / $0.150 131K โ€” 92.7%
SiliconFlow $0.075 $0.040 / $0.180 131K fp8 97.6%
Together $0.087 $0.050 / $0.200 131K โ€” 97.5%
Amazon Bedrock (eu-west-1) $0.090 $0.070 / $0.150 131K โ€” โ€”
Amazon Bedrock $0.090 $0.070 / $0.150 131K โ€” 99.8%
Google (us-central1) $0.115 $0.070 / $0.250 131K โ€” 58.1%
Fireworks $0.128 $0.070 / $0.300 131K โ€” โ€”
Groq $0.131 $0.075 / $0.300 131K โ€” 99.4%
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 โ€” the score above is not host-specific.
Compare gpt-oss-20b
FAQ
How much does gpt-oss-20b cost?
gpt-oss-20b costs $0.030 per million input tokens and $0.130 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.055 per million tokens.
What is gpt-oss-20b's context window?
gpt-oss-20b accepts up to 131,072 tokens of context, and can generate up to 131,072 output tokens.