token.app › Compare › Voxtral Small 24B 2507 vs DeepSeek V4 Flash 0731

Voxtral Small 24B 2507 vs DeepSeek V4 Flash 0731

DeepSeek V4 Flash 0731 is 1.3× cheaper per token on a blended 3:1 basis. DeepSeek V4 Flash 0731 accepts the larger context window (1.0M).

Head to head
Blended price uses a 3:1 input:output mix
Voxtral Small 24B 2507DeepSeek V4 Flash 0731
ProviderMistral AIDeepSeek
Input $/1M$0.100$0.090
Output $/1M$0.300$0.180
Blended $/1M$0.150$0.113
Context window32K1.0M
Max output384K
ReleasedOct 2025Jul 2026
LicenceOpen weightsOpen weights
Choose Voxtral Small 24B 2507 if…
  • No clear advantage on the data we hold.
Choose DeepSeek V4 Flash 0731 if…
  • Costs less per token — $0.113 vs $0.150 blended
  • Larger context window (1.0M)
  • Newer model (released Jul 2026)
FAQ
Which is better, Voxtral Small 24B 2507 or DeepSeek V4 Flash 0731?
We do not hold independently-run benchmark scores covering both Voxtral Small 24B 2507 and DeepSeek V4 Flash 0731, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Voxtral Small 24B 2507 cheaper than DeepSeek V4 Flash 0731?
DeepSeek V4 Flash 0731 is cheaper: $0.113 versus $0.150 per million tokens blended at 3:1 input:output. Input alone: $0.100 vs $0.090. Output alone: $0.300 vs $0.180.
What context windows do Voxtral Small 24B 2507 and DeepSeek V4 Flash 0731 have?
Voxtral Small 24B 2507: 32,000 tokens. DeepSeek V4 Flash 0731: 1,048,576 tokens.
Who makes Voxtral Small 24B 2507 and DeepSeek V4 Flash 0731?
Voxtral Small 24B 2507 is made by Mistral AI. DeepSeek V4 Flash 0731 is made by DeepSeek.
Related comparisons