token.app › Compare › Inkling Small vs Qwen3.7 Flash
Inkling Small vs Qwen3.7 Flash
Qwen3.7 Flash is 12× cheaper per token on a blended 3:1 basis. Qwen3.7 Flash accepts the larger context window (1M).
Head to head
Blended price uses a 3:1 input:output mix
| Inkling Small | Qwen3.7 Flash | |
|---|---|---|
| Provider | Thinkingmachines | Alibaba |
| Input $/1M | $0.450 | $0.030 |
| Output $/1M | $1.20 | $0.130 |
| Blended $/1M | $0.637 | $0.055 |
| Context window | 524K | 1M |
| Max output | 262K | 66K |
| Released | Jul 2026 | Jul 2026 |
| Licence | Open weights | Proprietary |
Choose Inkling Small if…
- Open weights — self-hostable
- Newer model (released Jul 2026)
Choose Qwen3.7 Flash if…
- Costs less per token — $0.055 vs $0.637 blended
- Larger context window (1M)
FAQ
Which is better, Inkling Small or Qwen3.7 Flash?
We do not hold independently-run benchmark scores covering both Inkling Small and Qwen3.7 Flash, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Inkling Small cheaper than Qwen3.7 Flash?
Qwen3.7 Flash is cheaper: $0.055 versus $0.637 per million tokens blended at 3:1 input:output. Input alone: $0.450 vs $0.030. Output alone: $1.20 vs $0.130.
What context windows do Inkling Small and Qwen3.7 Flash have?
Inkling Small: 524,288 tokens. Qwen3.7 Flash: 1,000,000 tokens.
Who makes Inkling Small and Qwen3.7 Flash?
Inkling Small is made by Thinkingmachines. Qwen3.7 Flash is made by Alibaba.