token.app › Compare › Gemini 3.1 Flash Lite vs Inkling Small

Gemini 3.1 Flash Lite vs Inkling Small

Gemini 3.1 Flash Lite is 1.1× cheaper per token on a blended 3:1 basis. Gemini 3.1 Flash Lite accepts the larger context window (1.0M).

Head to head
Blended price uses a 3:1 input:output mix
Gemini 3.1 Flash LiteInkling Small
ProviderGoogleThinkingmachines
Input $/1M$0.250$0.450
Output $/1M$1.50$1.20
Blended $/1M$0.563$0.637
Context window1.0M524K
Max output66K262K
ReleasedMay 2026Jul 2026
LicenceProprietaryOpen weights
Choose Gemini 3.1 Flash Lite if…
  • Costs less per token — $0.563 vs $0.637 blended
  • Larger context window (1.0M)
Choose Inkling Small if…
  • Open weights — self-hostable
  • Newer model (released Jul 2026)
FAQ
Which is better, Gemini 3.1 Flash Lite or Inkling Small?
We do not hold independently-run benchmark scores covering both Gemini 3.1 Flash Lite and Inkling Small, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Gemini 3.1 Flash Lite cheaper than Inkling Small?
Gemini 3.1 Flash Lite is cheaper: $0.563 versus $0.637 per million tokens blended at 3:1 input:output. Input alone: $0.250 vs $0.450. Output alone: $1.50 vs $1.20.
What context windows do Gemini 3.1 Flash Lite and Inkling Small have?
Gemini 3.1 Flash Lite: 1,048,576 tokens. Inkling Small: 524,288 tokens.
Who makes Gemini 3.1 Flash Lite and Inkling Small?
Gemini 3.1 Flash Lite is made by Google. Inkling Small is made by Thinkingmachines.
Related comparisons