Gemma 3 27B pricing & benchmarks
Gemma 3 27B is available from Google at $0.080 per million input tokens and $0.450 per million output tokens ($0.172 blended at 3:1). It accepts up to 262,144 tokens of context. Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
Pricing
Per million tokens ยท OpenRouter
Input$0.080
Output$0.450
Blended (3:1)$0.172
Specification
google/gemma-3-27b-it
Context window262K
Max output131K
ReleasedMar 2025
LicenceOpen weights
Modalitiesin:text, in:image, out:text
Benchmarks
Independently run โ not vendor self-reports. Every score links its source.
| Benchmark | Score | Config | Run | Source |
|---|---|---|---|---|
| GPQA Diamond | 43.9% ยฑ3.5 | โ | 2026-08-06 | Independently run ยท Epoch AI |
| OTIS Mock AIME | 22.2% ยฑ5.4 | โ | 2026-08-06 | Independently run ยท Epoch AI |
Scores published by Epoch AI, 'AI Benchmarking Hub', used under CC BY 4.0.
Cheaper models that score at least as well on GPQA Diamond
Strictly lower blended price AND an equal-or-higher GPQA Diamond score. One benchmark is one dimension โ a model that wins here may still be weaker at your specific task.
| Model | Provider | Blended $/1M | GPQA Diamond | |
|---|---|---|---|---|
| Mistral Small 3 | Mistral AI | $0.058 (โ67%) | 45.3% | Compare โ |
| gpt-oss-120b | OpenAI | $0.070 (โ59%) | 75.8% | Compare โ |
| Phi 4 | Microsoft | $0.088 (โ49%) | 56.1% | Compare โ |
| DeepSeek V4 Flash 0731 | DeepSeek | $0.113 (โ35%) | 91.0% | Compare โ |
| GPT-5 Nano | OpenAI | $0.138 (โ20%) | 69.4% | Compare โ |
Where to run Gemma 3 27B
5 hosts ยท 2.3ร between cheapest and dearest. Blended at 3:1. โ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
| Host | Blended $/1M | In / Out | Context | Weights | Uptime 24h |
|---|---|---|---|---|---|
| DeepInfra โ 131K context, not 262K |
$0.100 | $0.080 / $0.160 | 131K | fp8 | 98.5% |
| Novita โ 98K context, not 262K |
$0.139 | $0.119 / $0.200 | 98K | bf16 | 81.4% |
| Nebius โ 110K context, not 262K |
$0.150 | $0.100 / $0.300 | 110K | fp8 | 87.7% |
| Parasail โ 131K context, not 262K |
$0.172 | $0.080 / $0.450 | 131K | fp8 | 98.0% |
| Phala | $0.227 | $0.150 / $0.460 | 262K | โ | 83.8% |
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 โ the score above is not host-specific.
Compare Gemma 3 27B
FAQ
How much does Gemma 3 27B cost?
Gemma 3 27B costs $0.080 per million input tokens and $0.450 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.172 per million tokens.
What is Gemma 3 27B's context window?
Gemma 3 27B accepts up to 262,144 tokens of context, and can generate up to 131,072 output tokens.
How good is Gemma 3 27B on benchmarks?
Gemma 3 27B scores 43.9% on GPQA Diamond, independently run and published by Epoch AI. 2 benchmark results are listed on this page, each with its run date and source.
Is there a cheaper model as good as Gemma 3 27B?
Yes โ 5 models in our catalogue cost less per token than Gemma 3 27B and score at least as high on the same benchmark. The cheapest is Mistral Small 3 at $0.058 per million tokens blended.