token.app › Compare › Gemini 2.5 Flash Lite vs Inkling Small

Gemini 2.5 Flash Lite vs Inkling Small

Gemini 2.5 Flash Lite is 3.6× cheaper per token on a blended 3:1 basis. Gemini 2.5 Flash Lite accepts the larger context window (1.0M).

Head to head
Blended price uses a 3:1 input:output mix
Gemini 2.5 Flash LiteInkling Small
ProviderGoogleThinkingmachines
Input $/1M$0.100$0.450
Output $/1M$0.400$1.20
Blended $/1M$0.175$0.637
Context window1.0M524K
Max output66K262K
ReleasedJul 2025Jul 2026
LicenceProprietaryOpen weights
Choose Gemini 2.5 Flash Lite if…
  • Costs less per token — $0.175 vs $0.637 blended
  • Larger context window (1.0M)
Choose Inkling Small if…
  • Open weights — self-hostable
  • Newer model (released Jul 2026)
FAQ
Which is better, Gemini 2.5 Flash Lite or Inkling Small?
We do not hold independently-run benchmark scores covering both Gemini 2.5 Flash Lite and Inkling Small, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Gemini 2.5 Flash Lite cheaper than Inkling Small?
Gemini 2.5 Flash Lite is cheaper: $0.175 versus $0.637 per million tokens blended at 3:1 input:output. Input alone: $0.100 vs $0.450. Output alone: $0.400 vs $1.20.
What context windows do Gemini 2.5 Flash Lite and Inkling Small have?
Gemini 2.5 Flash Lite: 1,048,576 tokens. Inkling Small: 524,288 tokens.
Who makes Gemini 2.5 Flash Lite and Inkling Small?
Gemini 2.5 Flash Lite is made by Google. Inkling Small is made by Thinkingmachines.
Related comparisons