Compare · ModelsLive · 2 picked · head to head
phi-3-mini 3.8B vs Phi 4
Side by side · benchmarks, pricing, and signals you can act on.
Winner summary
phi-3-mini 3.8B wins on 1/2 benchmarks
phi-3-mini 3.8B wins 1 of 2 shared benchmarks. Leads in knowledge.
Category leads
knowledge·phi-3-mini 3.8B
Hype vs Reality
Attention vs performance
phi-3-mini 3.8B
#86 by perf·#20 by attention
Phi 4
#196 by perf·#20 by attention
Vendor risk
Who is behind the model
Microsoft
$3.84T·Big Tech
Microsoft
$3.84T·Big Tech
Head to head
2 benchmarks · 2 models
phi-3-mini 3.8BPhi 4
Chess Puzzles
Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities.
phi-3-mini 3.8B
0.0
Phi 4
0.0
MMLU
Phi 4 leads by +21.3
Massive Multitask Language Understanding · 57 subjects spanning STEM, humanities, social sciences, and more. The standard benchmark for broad knowledge.
phi-3-mini 3.8B
58.4
Phi 4
79.7
Full benchmark table
| Benchmark | phi-3-mini 3.8B | Phi 4 |
|---|---|---|
Chess Puzzles Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities. | 0.0 | 0.0 |
MMLU Massive Multitask Language Understanding · 57 subjects spanning STEM, humanities, social sciences, and more. The standard benchmark for broad knowledge. | 58.4 | 79.7 |
Pricing · per 1M tokens · projected $/mo at 10M tokens
| Model | Input | Output | Context | Projected $/mo |
|---|---|---|---|---|
| — | — | — | — | |
| $0.07 | $0.14 | 16K tokens (~8 books) | $0.88 |