token.app › Compare › Gemini 3.5 Flash vs Inkling Small

Gemini 3.5 Flash vs Inkling Small

Inkling Small is 5.3× cheaper per token on a blended 3:1 basis. Gemini 3.5 Flash accepts the larger context window (1.0M).

Head to head
Blended price uses a 3:1 input:output mix
Gemini 3.5 FlashInkling Small
ProviderGoogleThinkingmachines
Input $/1M$1.50$0.450
Output $/1M$9.00$1.20
Blended $/1M$3.38$0.637
Context window1.0M524K
Max output66K262K
ReleasedMay 2026Jul 2026
LicenceProprietaryOpen weights
Choose Gemini 3.5 Flash if…
  • Larger context window (1.0M)
Choose Inkling Small if…
  • Costs less per token — $0.637 vs $3.38 blended
  • Open weights — self-hostable
  • Newer model (released Jul 2026)
FAQ
Which is better, Gemini 3.5 Flash or Inkling Small?
We do not hold independently-run benchmark scores covering both Gemini 3.5 Flash and Inkling Small, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Gemini 3.5 Flash cheaper than Inkling Small?
Inkling Small is cheaper: $0.637 versus $3.38 per million tokens blended at 3:1 input:output. Input alone: $1.50 vs $0.450. Output alone: $9.00 vs $1.20.
What context windows do Gemini 3.5 Flash and Inkling Small have?
Gemini 3.5 Flash: 1,048,576 tokens. Inkling Small: 524,288 tokens.
Who makes Gemini 3.5 Flash and Inkling Small?
Gemini 3.5 Flash is made by Google. Inkling Small is made by Thinkingmachines.
Related comparisons