token.app › Compare › Gemini 3.6 Flash vs Inkling Small
Gemini 3.6 Flash vs Inkling Small
Inkling Small is 4.7× cheaper per token on a blended 3:1 basis. Gemini 3.6 Flash accepts the larger context window (1.0M).
Head to head
Blended price uses a 3:1 input:output mix
| Gemini 3.6 Flash | Inkling Small | |
|---|---|---|
| Provider | Thinkingmachines | |
| Input $/1M | $1.50 | $0.450 |
| Output $/1M | $7.50 | $1.20 |
| Blended $/1M | $3.00 | $0.637 |
| Context window | 1.0M | 524K |
| Max output | 66K | 262K |
| Released | Jul 2026 | Jul 2026 |
| Licence | Proprietary | Open weights |
Choose Gemini 3.6 Flash if…
- Larger context window (1.0M)
Choose Inkling Small if…
- Costs less per token — $0.637 vs $3.00 blended
- Open weights — self-hostable
- Newer model (released Jul 2026)
FAQ
Which is better, Gemini 3.6 Flash or Inkling Small?
We do not hold independently-run benchmark scores covering both Gemini 3.6 Flash and Inkling Small, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Gemini 3.6 Flash cheaper than Inkling Small?
Inkling Small is cheaper: $0.637 versus $3.00 per million tokens blended at 3:1 input:output. Input alone: $1.50 vs $0.450. Output alone: $7.50 vs $1.20.
What context windows do Gemini 3.6 Flash and Inkling Small have?
Gemini 3.6 Flash: 1,048,576 tokens. Inkling Small: 524,288 tokens.
Who makes Gemini 3.6 Flash and Inkling Small?
Gemini 3.6 Flash is made by Google. Inkling Small is made by Thinkingmachines.
Related comparisons