Gemini 3.5 Flash pricing & benchmarks
Gemini 3.5 Flash is available from Google at $1.50 per million input tokens and $9.00 per million output tokens ($3.38 blended at 3:1). It accepts up to 1,048,576 tokens of context. Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Pricing
Per million tokens · OpenRouter
Input$1.50
Output$9.00
Blended (3:1)$3.38
Per image<$0.01
Specification
google/gemini-3.5-flash
Context window1.0M
Max output66K
ReleasedMay 2026
LicenceProprietary
Modalitiesin:text, in:image, in:file, in:audio, in:video, out:text
Benchmarks
Independently run — not vendor self-reports. Every score links its source.
| Benchmark | Score | Config | Run | Source |
|---|---|---|---|---|
| GPQA Diamond | 92.8% ±1.6 | high effort | 2026-05-22 | Independently run · Epoch AI |
| SWE-bench Verified | 79.3% ±1.8 | high effort | 2026-06-01 | Independently run · Epoch AI |
| FrontierMath T1–3 | 62.8% ±2.9 | high effort | 2026-06-10 | Independently run · Epoch AI |
| FrontierMath T4 | 26.8% ±7.0 | high effort | 2026-06-10 | Independently run · Epoch AI |
| SimpleQA Verified | 68.4% ±1.5 | high effort | 2026-05-27 | Independently run · Epoch AI |
| OTIS Mock AIME | 95.6% ±2.7 | high effort | 2026-05-25 | Independently run · Epoch AI |
Scores published by Epoch AI, 'AI Benchmarking Hub', used under CC BY 4.0.
Cheaper models that score at least as well on GPQA Diamond
Strictly lower blended price AND an equal-or-higher GPQA Diamond score. One benchmark is one dimension — a model that wins here may still be weaker at your specific task.
| Model | Provider | Blended $/1M | GPQA Diamond | |
|---|---|---|---|---|
| GPT-5.6 Terra | OpenAI | $2.25 (−33%) | 93.3% | Compare → |
| Gemini 3.6 Flash | $3.00 (−11%) | 94.1% | Compare → | |
| Grok 4.5 | xAI | $3.00 (−11%) | 93.4% | Compare → |
Where to run Gemini 3.5 Flash
7 options across 2 providers · 3.6× between cheapest and dearest. Blended at 3:1. ⚠ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
| Host | Blended $/1M | In / Out | Context | Uptime 24h |
|---|---|---|---|---|
| Google (global/flex) | $1.69 | $0.750 / $4.50 | 1.0M | 99.0% |
| Google AI Studio (flex) | $1.69 | $0.750 / $4.50 | 1.0M | 99.9% |
| Google (global) | $3.38 | $1.50 / $9.00 | 1.0M | 99.0% |
| Google AI Studio | $3.38 | $1.50 / $9.00 | 1.0M | 99.9% |
| Google (us) | $3.71 | $1.65 / $9.90 | 1.0M | 100.0% |
| Google (global/priority) | $6.08 | $2.70 / $16.20 | 1.0M | 99.0% |
| Google AI Studio (priority) | $6.08 | $2.70 / $16.20 | 1.0M | 99.9% |
Host prices and uptime from OpenRouter, synced 2026-08-08.
Subscriptions that include Gemini 3.5 Flash
Consumer plans whose published model list names this model
Compare Gemini 3.5 Flash
FAQ
How much does Gemini 3.5 Flash cost?
Gemini 3.5 Flash costs $1.50 per million input tokens and $9.00 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $3.38 per million tokens.
What is Gemini 3.5 Flash's context window?
Gemini 3.5 Flash accepts up to 1,048,576 tokens of context, and can generate up to 65,536 output tokens.
How good is Gemini 3.5 Flash on benchmarks?
Gemini 3.5 Flash scores 92.8% on GPQA Diamond, independently run and published by Epoch AI. 6 benchmark results are listed on this page, each with its run date and source.
Is there a cheaper model as good as Gemini 3.5 Flash?
Yes — 3 models in our catalogue cost less per token than Gemini 3.5 Flash and score at least as high on the same benchmark. The cheapest is GPT-5.6 Terra at $2.25 per million tokens blended.