token.app › Compare › Gemini 3.1 Flash Lite (batch) vs Inkling Small

Gemini 3.1 Flash Lite (batch) vs Inkling Small

Gemini 3.1 Flash Lite (batch) is 2.3× cheaper per token on a blended 3:1 basis. Gemini 3.1 Flash Lite (batch) accepts the larger context window (1.0M).

Head to head
Blended price uses a 3:1 input:output mix
Gemini 3.1 Flash Lite (batch)Inkling Small
ProviderGoogleThinkingmachines
Input $/1M$0.125$0.450
Output $/1M$0.750$1.20
Blended $/1M$0.281$0.637
Context window1.0M524K
Max output66K262K
ReleasedMay 2026Jul 2026
LicenceProprietaryOpen weights
Choose Gemini 3.1 Flash Lite (batch) if…
  • Costs less per token — $0.281 vs $0.637 blended
  • Larger context window (1.0M)
Choose Inkling Small if…
  • Open weights — self-hostable
  • Newer model (released Jul 2026)
FAQ
Which is better, Gemini 3.1 Flash Lite (batch) or Inkling Small?
We do not hold independently-run benchmark scores covering both Gemini 3.1 Flash Lite (batch) and Inkling Small, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Gemini 3.1 Flash Lite (batch) cheaper than Inkling Small?
Gemini 3.1 Flash Lite (batch) is cheaper: $0.281 versus $0.637 per million tokens blended at 3:1 input:output. Input alone: $0.125 vs $0.450. Output alone: $0.750 vs $1.20.
What context windows do Gemini 3.1 Flash Lite (batch) and Inkling Small have?
Gemini 3.1 Flash Lite (batch): 1,048,576 tokens. Inkling Small: 524,288 tokens.
Who makes Gemini 3.1 Flash Lite (batch) and Inkling Small?
Gemini 3.1 Flash Lite (batch) is made by Google. Inkling Small is made by Thinkingmachines.
Related comparisons