Gemma 4 26B A4B pricing & benchmarks
Gemma 4 26B A4B is available from Google at $0.070 per million input tokens and $0.340 per million output tokens ($0.138 blended at 3:1). It accepts up to 262,144 tokens of context. Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference โ delivering near-31B quality at...
Pricing
Per million tokens ยท OpenRouter
Input$0.070
Output$0.340
Blended (3:1)$0.138
Specification
google/gemma-4-26b-a4b-it
Context window262K
Max output16K
ReleasedApr 2026
LicenceOpen weights
Modalitiesin:text, in:image, in:video, out:text
No independently-run benchmark scores are published for Gemma 4 26B A4B yet. We show nothing rather than reprinting an unverified figure.
Where to run Gemma 4 26B A4B
9 hosts ยท 1.9ร between cheapest and dearest. Blended at 3:1. โ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
| Host | Blended $/1M | In / Out | Context | Weights | Uptime 24h |
|---|---|---|---|---|---|
| DeepInfra | $0.138 | $0.070 / $0.340 | 262K | fp8 | 99.6% |
| Cloudflare | $0.150 | $0.100 / $0.300 | 256K | โ | 99.9% |
| NextBit | $0.190 | $0.120 / $0.400 | 262K | bf16 | 96.1% |
| SiliconFlow | $0.190 | $0.120 / $0.400 | 262K | fp8 | 99.3% |
| Novita | $0.198 | $0.130 / $0.400 | 262K | bf16 | 99.6% |
| Ionstream | $0.198 | $0.130 / $0.400 | 262K | bf16 | 98.7% |
| Parasail | $0.198 | $0.130 / $0.400 | 262K | bf16 | 97.2% |
| Venice | $0.198 | $0.130 / $0.400 | 256K | bf16 | 97.8% |
| Google (global) | $0.262 | $0.150 / $0.600 | 262K | โ | 97.2% |
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 โ the score above is not host-specific.
Compare Gemma 4 26B A4B
FAQ
How much does Gemma 4 26B A4B cost?
Gemma 4 26B A4B costs $0.070 per million input tokens and $0.340 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.138 per million tokens.
What is Gemma 4 26B A4B 's context window?
Gemma 4 26B A4B accepts up to 262,144 tokens of context, and can generate up to 16,384 output tokens.