token.app › Compare › Qwen3 30B A3B Thinking 2507 vs Gemini 3.6 Flash

Qwen3 30B A3B Thinking 2507 vs Gemini 3.6 Flash

Qwen3 30B A3B Thinking 2507 is 4.0× cheaper per token on a blended 3:1 basis. Gemini 3.6 Flash accepts the larger context window (1.0M).

Head to head
Blended price uses a 3:1 input:output mix
Qwen3 30B A3B Thinking 2507Gemini 3.6 Flash
ProviderAlibabaGoogle
Input $/1M$0.200$1.50
Output $/1M$2.40$7.50
Blended $/1M$0.750$3.00
Context window82K1.0M
Max output33K66K
ReleasedAug 2025Jul 2026
LicenceOpen weightsProprietary
Choose Qwen3 30B A3B Thinking 2507 if…
  • Costs less per token — $0.750 vs $3.00 blended
  • Open weights — self-hostable
Choose Gemini 3.6 Flash if…
  • Larger context window (1.0M)
  • Newer model (released Jul 2026)
FAQ
Which is better, Qwen3 30B A3B Thinking 2507 or Gemini 3.6 Flash?
We do not hold independently-run benchmark scores covering both Qwen3 30B A3B Thinking 2507 and Gemini 3.6 Flash, so we make no quality claim. Their pricing and specifications are compared on this page.
Is Qwen3 30B A3B Thinking 2507 cheaper than Gemini 3.6 Flash?
Qwen3 30B A3B Thinking 2507 is cheaper: $0.750 versus $3.00 per million tokens blended at 3:1 input:output. Input alone: $0.200 vs $1.50. Output alone: $2.40 vs $7.50.
What context windows do Qwen3 30B A3B Thinking 2507 and Gemini 3.6 Flash have?
Qwen3 30B A3B Thinking 2507: 81,920 tokens. Gemini 3.6 Flash: 1,048,576 tokens.
Who makes Qwen3 30B A3B Thinking 2507 and Gemini 3.6 Flash?
Qwen3 30B A3B Thinking 2507 is made by Alibaba. Gemini 3.6 Flash is made by Google.
Related comparisons