Compare · ModelsLive · 2 picked · head to head
Kimi K2 Thinking vs Qwen2.5 32B Instruct
Side by side · benchmarks, pricing, and signals you can act on.
Winner summary
Kimi K2 Thinking wins on 3/3 benchmarks
Kimi K2 Thinking wins 3 of 3 shared benchmarks. Leads in knowledge · math.
Category leads
knowledge·Kimi K2 Thinkingmath·Kimi K2 Thinking
Hype vs Reality
Attention vs performance
Kimi K2 Thinking
#103 by perf·#17 by attention
Qwen2.5 32B Instruct
#222 by perf·#2 by attention
Best value
Kimi K2 Thinking
Kimi K2 Thinking
34.0 pts/$
$1.55/M
Qwen2.5 32B Instruct
n/a
no price
Vendor risk
Who is behind the model
Moonshot AI
$18.0B·Tier 1
Alibaba (Qwen)
$293.0B·Tier 1
Head to head
3 benchmarks · 2 models
Kimi K2 ThinkingQwen2.5 32B Instruct
Chess Puzzles
Kimi K2 Thinking leads by +15.8
Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities.
Kimi K2 Thinking
15.8
Qwen2.5 32B Instruct
0.0
GPQA diamond
Kimi K2 Thinking leads by +50.8
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
Kimi K2 Thinking
79.0
Qwen2.5 32B Instruct
28.1
OTIS Mock AIME 2024-2025
Kimi K2 Thinking leads by +75.8
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
Kimi K2 Thinking
83.0
Qwen2.5 32B Instruct
7.3
Full benchmark table
| Benchmark | Kimi K2 Thinking | Qwen2.5 32B Instruct |
|---|---|---|
Chess Puzzles Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities. | 15.8 | 0.0 |
GPQA diamond Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs. | 79.0 | 28.1 |
OTIS Mock AIME 2024-2025 OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills. | 83.0 | 7.3 |
Pricing · per 1M tokens · projected $/mo at 10M tokens
| Model | Input | Output | Context | Projected $/mo |
|---|---|---|---|---|
| $0.60 | $2.50 | 262K tokens (~131 books) | $10.75 | |
| — | — | — | — |