token.app › Compare › Gemini 3.5 Flash (batch) vs Inkling Small
Gemini 3.5 Flash (batch) vs Inkling Small
Inkling Small is 2.6× cheaper per token on a blended 3:1 basis. Gemini 3.5 Flash (batch) accepts the larger context window (1.0M).
Head to head
Blended price uses a 3:1 input:output mix
| Gemini 3.5 Flash (batch) | Inkling Small | |
|---|---|---|
| Provider | Thinkingmachines | |
| Input $/1M | $0.750 | $0.450 |
| Output $/1M | $4.50 | $1.20 |
| Blended $/1M | $1.69 | $0.637 |
| Context window | 1.0M | 524K |
| Max output | 66K | 262K |
| Released | May 2026 | Jul 2026 |
| Licence | Proprietary | Open weights |
Choose Gemini 3.5 Flash (batch) if…
- Larger context window (1.0M)
Choose Inkling Small if…
- Costs less per token — $0.637 vs $1.69 blended
- Open weights — self-hostable
- Newer model (released Jul 2026)
FAQ
Which is better, Gemini 3.5 Flash (batch) or Inkling Small?
We do not hold independently-run benchmark scores covering both Gemini 3.5 Flash (batch) and Inkling Small, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Gemini 3.5 Flash (batch) cheaper than Inkling Small?
Inkling Small is cheaper: $0.637 versus $1.69 per million tokens blended at 3:1 input:output. Input alone: $0.750 vs $0.450. Output alone: $4.50 vs $1.20.
What context windows do Gemini 3.5 Flash (batch) and Inkling Small have?
Gemini 3.5 Flash (batch): 1,048,576 tokens. Inkling Small: 524,288 tokens.
Who makes Gemini 3.5 Flash (batch) and Inkling Small?
Gemini 3.5 Flash (batch) is made by Google. Inkling Small is made by Thinkingmachines.
Related comparisons
Gemini 3.5 Flash (batch) full details
Inkling Small full details
Gemini 3.5 Flash (batch) vs Ling 3.0 Tiny (free)
Gemini 3.5 Flash (batch) vs Muse Spark 1.2
Gemini 3.5 Flash (batch) vs Qwen3.8 Max
Gemini 3.5 Flash (batch) vs DeepSeek V4 Flash Latest
Gemini 3.5 Flash (batch) vs DeepSeek V4 Flash 0731
Gemini 3.5 Flash (batch) vs Qwen3.7 Flash