token.app › Compare › DeepSeek V4 Flash 0731 vs Inkling Small

DeepSeek V4 Flash 0731 vs Inkling Small

DeepSeek V4 Flash 0731 is 5.7× cheaper per token on a blended 3:1 basis. DeepSeek V4 Flash 0731 accepts the larger context window (1.0M).

Head to head
Blended price uses a 3:1 input:output mix
DeepSeek V4 Flash 0731Inkling Small
ProviderDeepSeekThinkingmachines
Input $/1M$0.090$0.450
Output $/1M$0.180$1.20
Blended $/1M$0.113$0.637
Context window1.0M524K
Max output384K262K
ReleasedJul 2026Jul 2026
LicenceOpen weightsOpen weights
Choose DeepSeek V4 Flash 0731 if…
  • Costs less per token — $0.113 vs $0.637 blended
  • Larger context window (1.0M)
  • Newer model (released Jul 2026)
Choose Inkling Small if…
  • No clear advantage on the data we hold.
FAQ
Which is better, DeepSeek V4 Flash 0731 or Inkling Small?
We do not hold independently-run benchmark scores covering both DeepSeek V4 Flash 0731 and Inkling Small, so we make no quality claim. Their pricing and specifications are compared on this page.
Is DeepSeek V4 Flash 0731 cheaper than Inkling Small?
DeepSeek V4 Flash 0731 is cheaper: $0.113 versus $0.637 per million tokens blended at 3:1 input:output. Input alone: $0.090 vs $0.450. Output alone: $0.180 vs $1.20.
What context windows do DeepSeek V4 Flash 0731 and Inkling Small have?
DeepSeek V4 Flash 0731: 1,048,576 tokens. Inkling Small: 524,288 tokens.
Who makes DeepSeek V4 Flash 0731 and Inkling Small?
DeepSeek V4 Flash 0731 is made by DeepSeek. Inkling Small is made by Thinkingmachines.
Related comparisons