token.app › Compare › DeepSeek V4 Flash Latest vs Inkling Small

DeepSeek V4 Flash Latest vs Inkling Small

DeepSeek V4 Flash Latest is 5.7× cheaper per token on a blended 3:1 basis. DeepSeek V4 Flash Latest accepts the larger context window (1.0M).

Head to head
Blended price uses a 3:1 input:output mix
DeepSeek V4 Flash LatestInkling Small
Provider~deepseekThinkingmachines
Input $/1M$0.090$0.450
Output $/1M$0.180$1.20
Blended $/1M$0.113$0.637
Context window1.0M524K
Max output384K262K
ReleasedAug 2026Jul 2026
LicenceProprietaryOpen weights
Choose DeepSeek V4 Flash Latest if…
  • Costs less per token — $0.113 vs $0.637 blended
  • Larger context window (1.0M)
  • Newer model (released Aug 2026)
Choose Inkling Small if…
  • Open weights — self-hostable
FAQ
Which is better, DeepSeek V4 Flash Latest or Inkling Small?
We do not hold independently-run benchmark scores covering both DeepSeek V4 Flash Latest and Inkling Small, so we make no quality claim. Their pricing and specifications are compared on this page.
Is DeepSeek V4 Flash Latest cheaper than Inkling Small?
DeepSeek V4 Flash Latest is cheaper: $0.113 versus $0.637 per million tokens blended at 3:1 input:output. Input alone: $0.090 vs $0.450. Output alone: $0.180 vs $1.20.
What context windows do DeepSeek V4 Flash Latest and Inkling Small have?
DeepSeek V4 Flash Latest: 1,048,576 tokens. Inkling Small: 524,288 tokens.
Who makes DeepSeek V4 Flash Latest and Inkling Small?
DeepSeek V4 Flash Latest is made by ~deepseek. Inkling Small is made by Thinkingmachines.
Related comparisons