Gemma 4 26B A4B (free) pricing & benchmarks
Gemma 4 26B A4B (free) is available from Google at Free per million input tokens and Free per million output tokens (Free blended at 3:1). It accepts up to 262,144 tokens of context. Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference โ delivering near-31B quality at...
Pricing
Per million tokens ยท OpenRouter
InputFree
OutputFree
Blended (3:1)Free
Specification
google/gemma-4-26b-a4b-it:free
Context window262K
Max output33K
ReleasedApr 2026
LicenceOpen weights
Modalitiesin:text, in:image, in:video, out:text
No independently-run benchmark scores are published for Gemma 4 26B A4B (free) yet. We show nothing rather than reprinting an unverified figure.
Where to run Gemma 4 26B A4B (free)
2 hosts ยท all priced within ~5% of each other. Blended at 3:1. โ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
| Host | Blended $/1M | In / Out | Context | Uptime 24h |
|---|---|---|---|---|
| Darkbloom โ 131K context, not 262K |
Free | Free / Free | 131K | 97.1% |
| Google AI Studio | Free | Free / Free | 262K | 99.3% |
Host prices and uptime from OpenRouter, synced 2026-08-08.
Compare Gemma 4 26B A4B (free)
Gemma 4 26B A4B (free) vs Qwen3.8 Max
Gemma 4 26B A4B (free) vs DeepSeek V4 Flash 0731
Gemma 4 26B A4B (free) vs Claude Opus 5
Gemma 4 26B A4B (free) vs Kimi K3
Gemma 4 26B A4B (free) vs GPT-5.6 Luna
Gemma 4 26B A4B (free) vs GPT-5.6 Terra
Gemma 4 26B A4B (free) vs GPT-5.6 Sol
Gemma 4 26B A4B (free) vs Grok 4.5
FAQ
How much does Gemma 4 26B A4B (free) cost?
Gemma 4 26B A4B (free) costs Free per million input tokens and Free per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to Free per million tokens.
What is Gemma 4 26B A4B (free)'s context window?
Gemma 4 26B A4B (free) accepts up to 262,144 tokens of context, and can generate up to 32,768 output tokens.