Compare · ModelsLive · 2 picked · head to head

Claude 2.1 vs Mistral Small 3.1 24B

Side by side · benchmarks, pricing, and signals you can act on.

Winner summary

Mistral Small 3.1 24B wins 3 of 3 shared benchmarks. Leads in general · knowledge · math.

Category leads
general·Mistral Small 3.1 24Bknowledge·Mistral Small 3.1 24Bmath·Mistral Small 3.1 24B
Hype vs Reality
Claude 2.1
#286 by perf·no signal
QUIET
Mistral Small 3.1 24B
#194 by perf·no signal
QUIET
Best value
Claude 2.1
n/a
no price
Mistral Small 3.1 24B
86.5 pts/$
$0.45/M
Vendor risk
Anthropic logo
Anthropic
$965.0B·Tier 1
Medium risk
Mistral AI logo
Mistral AI
$14.0B·Tier 1
Medium risk
Head to head
Claude 2.1Mistral Small 3.1 24B
Dtbench
Mistral Small 3.1 24B leads by +12.8
Claude 2.1
18.3
Mistral Small 3.1 24B
31.1
GPQA diamond
Mistral Small 3.1 24B leads by +19.4
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
Claude 2.1
10.6
Mistral Small 3.1 24B
30.0
OTIS Mock AIME 2024-2025
Mistral Small 3.1 24B leads by +3.9
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
Claude 2.1
1.9
Mistral Small 3.1 24B
5.7
Full benchmark table
BenchmarkClaude 2.1Mistral Small 3.1 24B
Dtbench
18.331.1
GPQA diamond
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
10.630.0
OTIS Mock AIME 2024-2025
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
1.95.7
Pricing · per 1M tokens · projected $/mo at 10M tokens
ModelInputOutputContextProjected $/mo
Anthropic logoClaude 2.1————
Mistral AI logoMistral Small 3.1 24B$0.35$0.56128K tokens (~64 books)$4.02