token.app › Compare › Gemini 2.5 Flash Lite vs Inkling Small
Gemini 2.5 Flash Lite vs Inkling Small
Gemini 2.5 Flash Lite is 3.6× cheaper per token on a blended 3:1 basis. Gemini 2.5 Flash Lite accepts the larger context window (1.0M).
Head to head
Blended price uses a 3:1 input:output mix
| Gemini 2.5 Flash Lite | Inkling Small | |
|---|---|---|
| Provider | Thinkingmachines | |
| Input $/1M | $0.100 | $0.450 |
| Output $/1M | $0.400 | $1.20 |
| Blended $/1M | $0.175 | $0.637 |
| Context window | 1.0M | 524K |
| Max output | 66K | 262K |
| Released | Jul 2025 | Jul 2026 |
| Licence | Proprietary | Open weights |
Choose Gemini 2.5 Flash Lite if…
- Costs less per token — $0.175 vs $0.637 blended
- Larger context window (1.0M)
Choose Inkling Small if…
- Open weights — self-hostable
- Newer model (released Jul 2026)
FAQ
Which is better, Gemini 2.5 Flash Lite or Inkling Small?
We do not hold independently-run benchmark scores covering both Gemini 2.5 Flash Lite and Inkling Small, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Gemini 2.5 Flash Lite cheaper than Inkling Small?
Gemini 2.5 Flash Lite is cheaper: $0.175 versus $0.637 per million tokens blended at 3:1 input:output. Input alone: $0.100 vs $0.450. Output alone: $0.400 vs $1.20.
What context windows do Gemini 2.5 Flash Lite and Inkling Small have?
Gemini 2.5 Flash Lite: 1,048,576 tokens. Inkling Small: 524,288 tokens.
Who makes Gemini 2.5 Flash Lite and Inkling Small?
Gemini 2.5 Flash Lite is made by Google. Inkling Small is made by Thinkingmachines.
Related comparisons
Gemini 2.5 Flash Lite full details
Inkling Small full details
Gemini 2.5 Flash Lite vs Ling 3.0 Tiny (free)
Gemini 2.5 Flash Lite vs Muse Spark 1.2
Gemini 2.5 Flash Lite vs Qwen3.8 Max
Gemini 2.5 Flash Lite vs DeepSeek V4 Flash Latest
Gemini 2.5 Flash Lite vs DeepSeek V4 Flash 0731
Gemini 2.5 Flash Lite vs Qwen3.7 Flash