token.app › Compare › Inkling Small vs Qwen3.7 Flash

Inkling Small vs Qwen3.7 Flash

Qwen3.7 Flash is 12× cheaper per token on a blended 3:1 basis. Qwen3.7 Flash accepts the larger context window (1M).

Head to head
Blended price uses a 3:1 input:output mix
Inkling SmallQwen3.7 Flash
ProviderThinkingmachinesAlibaba
Input $/1M$0.450$0.030
Output $/1M$1.20$0.130
Blended $/1M$0.637$0.055
Context window524K1M
Max output262K66K
ReleasedJul 2026Jul 2026
LicenceOpen weightsProprietary
Choose Inkling Small if…
  • Open weights — self-hostable
  • Newer model (released Jul 2026)
Choose Qwen3.7 Flash if…
  • Costs less per token — $0.055 vs $0.637 blended
  • Larger context window (1M)
FAQ
Which is better, Inkling Small or Qwen3.7 Flash?
We do not hold independently-run benchmark scores covering both Inkling Small and Qwen3.7 Flash, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Inkling Small cheaper than Qwen3.7 Flash?
Qwen3.7 Flash is cheaper: $0.055 versus $0.637 per million tokens blended at 3:1 input:output. Input alone: $0.450 vs $0.030. Output alone: $1.20 vs $0.130.
What context windows do Inkling Small and Qwen3.7 Flash have?
Inkling Small: 524,288 tokens. Qwen3.7 Flash: 1,000,000 tokens.
Who makes Inkling Small and Qwen3.7 Flash?
Inkling Small is made by Thinkingmachines. Qwen3.7 Flash is made by Alibaba.
Related comparisons