token.app › Compare › Gemini 3.5 Flash (batch) vs Inkling Small

Gemini 3.5 Flash (batch) vs Inkling Small

Inkling Small is 2.6× cheaper per token on a blended 3:1 basis. Gemini 3.5 Flash (batch) accepts the larger context window (1.0M).

Head to head
Blended price uses a 3:1 input:output mix
Gemini 3.5 Flash (batch)Inkling Small
ProviderGoogleThinkingmachines
Input $/1M$0.750$0.450
Output $/1M$4.50$1.20
Blended $/1M$1.69$0.637
Context window1.0M524K
Max output66K262K
ReleasedMay 2026Jul 2026
LicenceProprietaryOpen weights
Choose Gemini 3.5 Flash (batch) if…
  • Larger context window (1.0M)
Choose Inkling Small if…
  • Costs less per token — $0.637 vs $1.69 blended
  • Open weights — self-hostable
  • Newer model (released Jul 2026)
FAQ
Which is better, Gemini 3.5 Flash (batch) or Inkling Small?
We do not hold independently-run benchmark scores covering both Gemini 3.5 Flash (batch) and Inkling Small, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Gemini 3.5 Flash (batch) cheaper than Inkling Small?
Inkling Small is cheaper: $0.637 versus $1.69 per million tokens blended at 3:1 input:output. Input alone: $0.750 vs $0.450. Output alone: $4.50 vs $1.20.
What context windows do Gemini 3.5 Flash (batch) and Inkling Small have?
Gemini 3.5 Flash (batch): 1,048,576 tokens. Inkling Small: 524,288 tokens.
Who makes Gemini 3.5 Flash (batch) and Inkling Small?
Gemini 3.5 Flash (batch) is made by Google. Inkling Small is made by Thinkingmachines.
Related comparisons