token.appxAI › Grok 4.20

Grok 4.20 pricing & benchmarks

Grok 4.20 is available from xAI at $1.25 per million input tokens and $2.50 per million output tokens ($1.56 blended at 3:1). It accepts up to 2,000,000 tokens of context. Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...

Pricing
Per million tokens · OpenRouter
Input$1.25
Output$2.50
Blended (3:1)$1.56
Specification
x-ai/grok-4.20
Context window2M
Max output
ReleasedMar 2026
LicenceProprietary
Modalitiesin:text, in:image, in:file, out:text
Benchmarks
Independently run — not vendor self-reports. Every score links its source.
BenchmarkScoreConfigRunSource
GPQA Diamond 89.3% ±1.9 2026-07-13 Independently run · Epoch AI
FrontierMath T1–3 44.9% ±3.0 2026-07-13 Independently run · Epoch AI
FrontierMath T4 17.1% ±5.9 2026-07-13 Independently run · Epoch AI
SimpleQA Verified 35.1% ±1.5 2026-07-13 Independently run · Epoch AI
OTIS Mock AIME 92.2% ±3.3 2026-07-13 Independently run · Epoch AI
Scores published by Epoch AI, 'AI Benchmarking Hub', used under CC BY 4.0.
Cheaper models that score at least as well on GPQA Diamond
Strictly lower blended price AND an equal-or-higher GPQA Diamond score. One benchmark is one dimension — a model that wins here may still be weaker at your specific task.
ModelProviderBlended $/1MGPQA Diamond
DeepSeek V4 Flash 0731 DeepSeek $0.113 (−93%) 91.0% Compare →
GPT-5.6 Luna OpenAI $0.225 (−86%) 91.6% Compare →
GLM 5.2 Zhipu AI $0.387 (−75%) 91.9% Compare →
DeepSeek V4 Pro DeepSeek $0.544 (−65%) 90.9% Compare →
Kimi K2.6 Moonshot AI $1.04 (−33%) 90.8% Compare →
Where to run Grok 4.20
4 options across 1 provider · 2.0× between cheapest and dearest. Blended at 3:1. ⚠ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
HostBlended $/1MIn / OutContextUptime 24h
xAI
⚠ re-prices above 200K prompt tokens ($2.50/$5.00)
$1.56 $1.25 / $2.50 2M 99.9%
xAI (zdr)
⚠ re-prices above 200K prompt tokens ($2.50/$5.00)
$1.56 $1.25 / $2.50 2M 100.0%
xAI (priority)
⚠ re-prices above 200K prompt tokens ($5.00/$10.00)
$3.13 $2.50 / $5.00 2M 99.9%
xAI (zdr/priority)
⚠ re-prices above 200K prompt tokens ($5.00/$10.00)
$3.13 $2.50 / $5.00 2M 100.0%
Host prices and uptime from OpenRouter, synced 2026-08-08.
Compare Grok 4.20
FAQ
How much does Grok 4.20 cost?
Grok 4.20 costs $1.25 per million input tokens and $2.50 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $1.56 per million tokens.
What is Grok 4.20's context window?
Grok 4.20 accepts up to 2,000,000 tokens of context.
How good is Grok 4.20 on benchmarks?
Grok 4.20 scores 89.3% on GPQA Diamond, independently run and published by Epoch AI. 5 benchmark results are listed on this page, each with its run date and source.
Is there a cheaper model as good as Grok 4.20?
Yes — 5 models in our catalogue cost less per token than Grok 4.20 and score at least as high on the same benchmark. The cheapest is DeepSeek V4 Flash 0731 at $0.113 per million tokens blended.