token.app › Compare › Inkling Small vs Qwen3.8 Max
Inkling Small vs Qwen3.8 Max
Inkling Small is 4.7× cheaper per token on a blended 3:1 basis. Qwen3.8 Max accepts the larger context window (1M).
Head to head
Blended price uses a 3:1 input:output mix
| Inkling Small | Qwen3.8 Max | |
|---|---|---|
| Provider | Thinkingmachines | Alibaba |
| Input $/1M | $0.450 | $2.00 |
| Output $/1M | $1.20 | $6.00 |
| Blended $/1M | $0.637 | $3.00 |
| Context window | 524K | 1M |
| Max output | 262K | 131K |
| Released | Jul 2026 | Aug 2026 |
| Licence | Open weights | Proprietary |
Choose Inkling Small if…
- Costs less per token — $0.637 vs $3.00 blended
- Open weights — self-hostable
Choose Qwen3.8 Max if…
- Larger context window (1M)
- Newer model (released Aug 2026)
FAQ
Which is better, Inkling Small or Qwen3.8 Max?
We do not hold independently-run benchmark scores covering both Inkling Small and Qwen3.8 Max, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Inkling Small cheaper than Qwen3.8 Max?
Inkling Small is cheaper: $0.637 versus $3.00 per million tokens blended at 3:1 input:output. Input alone: $0.450 vs $2.00. Output alone: $1.20 vs $6.00.
What context windows do Inkling Small and Qwen3.8 Max have?
Inkling Small: 524,288 tokens. Qwen3.8 Max: 1,000,000 tokens.
Who makes Inkling Small and Qwen3.8 Max?
Inkling Small is made by Thinkingmachines. Qwen3.8 Max is made by Alibaba.