token.app › Compare › Gemini 2.5 Flash (batch) vs Inkling Small
Gemini 2.5 Flash (batch) vs Inkling Small
Gemini 2.5 Flash (batch) is 1.5× cheaper per token on a blended 3:1 basis. Gemini 2.5 Flash (batch) accepts the larger context window (1.0M).
Head to head
Blended price uses a 3:1 input:output mix
| Gemini 2.5 Flash (batch) | Inkling Small | |
|---|---|---|
| Provider | Thinkingmachines | |
| Input $/1M | $0.150 | $0.450 |
| Output $/1M | $1.25 | $1.20 |
| Blended $/1M | $0.425 | $0.637 |
| Context window | 1.0M | 524K |
| Max output | 66K | 262K |
| Released | Jun 2025 | Jul 2026 |
| Licence | Proprietary | Open weights |
Choose Gemini 2.5 Flash (batch) if…
- Costs less per token — $0.425 vs $0.637 blended
- Larger context window (1.0M)
Choose Inkling Small if…
- Open weights — self-hostable
- Newer model (released Jul 2026)
FAQ
Which is better, Gemini 2.5 Flash (batch) or Inkling Small?
We do not hold independently-run benchmark scores covering both Gemini 2.5 Flash (batch) and Inkling Small, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Gemini 2.5 Flash (batch) cheaper than Inkling Small?
Gemini 2.5 Flash (batch) is cheaper: $0.425 versus $0.637 per million tokens blended at 3:1 input:output. Input alone: $0.150 vs $0.450. Output alone: $1.25 vs $1.20.
What context windows do Gemini 2.5 Flash (batch) and Inkling Small have?
Gemini 2.5 Flash (batch): 1,048,576 tokens. Inkling Small: 524,288 tokens.
Who makes Gemini 2.5 Flash (batch) and Inkling Small?
Gemini 2.5 Flash (batch) is made by Google. Inkling Small is made by Thinkingmachines.
Related comparisons
Gemini 2.5 Flash (batch) full details
Inkling Small full details
Gemini 2.5 Flash (batch) vs Ling 3.0 Tiny (free)
Gemini 2.5 Flash (batch) vs Muse Spark 1.2
Gemini 2.5 Flash (batch) vs Qwen3.8 Max
Gemini 2.5 Flash (batch) vs DeepSeek V4 Flash Latest
Gemini 2.5 Flash (batch) vs DeepSeek V4 Flash 0731
Gemini 2.5 Flash (batch) vs Qwen3.7 Flash