token.app โ€บ Compare โ€บ Gemini 2.5 Flash vs Kimi K3

Gemini 2.5 Flash vs Kimi K3

Kimi K3 leads on 1 of 1 shared benchmarks. Gemini 2.5 Flash is 7.1ร— cheaper per token on a blended 3:1 basis.

Head to head
Blended price uses a 3:1 input:output mix
Gemini 2.5 FlashKimi K3
ProviderGoogleMoonshot AI
Input $/1M$0.300$3.00
Output $/1M$2.50$15.00
Blended $/1M$0.850$6.00
Context window1.0M1.0M
Max output66Kโ€”
ReleasedJun 2025Jul 2026
LicenceProprietaryOpen weights
OTIS Mock AIME 70.8% 97.2%
Benchmark scores independently run and published by Epoch AI, 'AI Benchmarking Hub', used under CC BY 4.0. Only benchmarks with a score for BOTH models are compared.
Choose Gemini 2.5 Flash ifโ€ฆ
  • Costs less per token โ€” $0.850 vs $6.00 blended
Choose Kimi K3 ifโ€ฆ
  • Higher OTIS Mock AIME (97.2% vs 70.8%)
  • Open weights โ€” self-hostable
  • Newer model (released Jul 2026)
FAQ
Which is better, Gemini 2.5 Flash or Kimi K3?
Across the 1 benchmark both models have independently-run scores for, Gemini 2.5 Flash leads on 0 and Kimi K3 leads on 1. Benchmarks are one input โ€” price and context window are on this page too.
Is Gemini 2.5 Flash cheaper than Kimi K3?
Gemini 2.5 Flash is cheaper: $0.850 versus $6.00 per million tokens blended at 3:1 input:output. Input alone: $0.300 vs $3.00. Output alone: $2.50 vs $15.00.
What context windows do Gemini 2.5 Flash and Kimi K3 have?
Gemini 2.5 Flash: 1,048,576 tokens. Kimi K3: 1,048,576 tokens.
Who makes Gemini 2.5 Flash and Kimi K3?
Gemini 2.5 Flash is made by Google. Kimi K3 is made by Moonshot AI.
Related comparisons