Compare · ModelsLive · 2 picked · head to head
Llama 3.1 405B vs Qwen 3.5 Plus (hosted 397B-A17B)
Side by side · benchmarks, pricing, and signals you can act on.
Winner summary
Qwen 3.5 Plus (hosted 397B-A17B) wins on 3/3 benchmarks
Qwen 3.5 Plus (hosted 397B-A17B) wins 3 of 3 shared benchmarks. Leads in general · knowledge · math.
Category leads
general·Qwen 3.5 Plus (hosted 397B-A17B)knowledge·Qwen 3.5 Plus (hosted 397B-A17B)math·Qwen 3.5 Plus (hosted 397B-A17B)
Hype vs Reality
Attention vs performance
Llama 3.1 405B
#206 by perf·no signal
Qwen 3.5 Plus (hosted 397B-A17B)
#228 by perf·#2 by attention
Best value
Pricing unknown
Llama 3.1 405B
n/a
no price
Qwen 3.5 Plus (hosted 397B-A17B)
n/a
no price
Vendor risk
Who is behind the model
Meta AI
$1.87T·Tier 1
Alibaba (Qwen)
$293.0B·Tier 1
Head to head
3 benchmarks · 2 models
Llama 3.1 405BQwen 3.5 Plus (hosted 397B-A17B)
Dtbench
Qwen 3.5 Plus (hosted 397B-A17B) leads by +31.9
Llama 3.1 405B
35.6
Qwen 3.5 Plus (hosted 397B-A17B)
67.5
GPQA diamond
Qwen 3.5 Plus (hosted 397B-A17B) leads by +45.3
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
Llama 3.1 405B
34.5
Qwen 3.5 Plus (hosted 397B-A17B)
79.8
OTIS Mock AIME 2024-2025
Qwen 3.5 Plus (hosted 397B-A17B) leads by +77.0
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
Llama 3.1 405B
9.6
Qwen 3.5 Plus (hosted 397B-A17B)
86.7
Full benchmark table
| Benchmark | Llama 3.1 405B | Qwen 3.5 Plus (hosted 397B-A17B) |
|---|---|---|
Dtbench | 35.6 | 67.5 |
GPQA diamond Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs. | 34.5 | 79.8 |
OTIS Mock AIME 2024-2025 OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills. | 9.6 | 86.7 |
Pricing · per 1M tokens · projected $/mo at 10M tokens
| Model | Input | Output | Context | Projected $/mo |
|---|---|---|---|---|
| — | — | — | — | |
| — | — | — | — |
People also compared