Compare · ModelsLive · 2 picked · head to head
GLM 5 vs Qwen3.5 397B A17B
Side by side · benchmarks, pricing, and signals you can act on.
Winner summary
GLM 5 wins on 7/12 benchmarks
GLM 5 wins 7 of 12 shared benchmarks. Leads in arena · math · language.
Category leads
agentic·Qwen3.5 397B A17Barena·GLM 5knowledge·Qwen3.5 397B A17Bmath·GLM 5language·GLM 5coding·GLM 5
Hype vs Reality
Attention vs performance
GLM 5
#79 by perf·#3 by attention
Qwen3.5 397B A17B
#48 by perf·#2 by attention
Best value
GLM 5
1.5x better value than Qwen3.5 397B A17B
GLM 5
43.7 pts/$
$1.26/M
Qwen3.5 397B A17B
29.6 pts/$
$2.02/M
Vendor risk
Who is behind the model
z-ai
private · undisclosed
Alibaba (Qwen)
$293.0B·Tier 1
Head to head
12 benchmarks · 2 models
GLM 5Qwen3.5 397B A17B
APEX-Agents
Qwen3.5 397B A17B leads by +7.7
APEX-Agents · evaluates AI agents on complex, multi-step tasks requiring planning, tool use, and autonomous decision-making in realistic environments.
GLM 5
17.2
Qwen3.5 397B A17B
24.9
Chatbot Arena Elo · Coding
GLM 5 leads by +34.6
GLM 5
1434.0
Qwen3.5 397B A17B
1399.4
Chatbot Arena Elo · Overall
GLM 5 leads by +15.8
GLM 5
1457.6
Qwen3.5 397B A17B
1441.8
Chess Puzzles
Qwen3.5 397B A17B leads by +3.2
Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities.
GLM 5
5.3
Qwen3.5 397B A17B
8.5
GPQA diamond
GLM 5 leads by +1.9
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
GLM 5
83.8
Qwen3.5 397B A17B
81.8
OpenCompass · AIME2025
GLM 5 leads by +3.5
GLM 5
95.8
Qwen3.5 397B A17B
92.3
OpenCompass · GPQA-Diamond
Qwen3.5 397B A17B leads by +3.1
GLM 5
85.3
Qwen3.5 397B A17B
88.4
OpenCompass · HLE
GLM 5 leads by +0.6
GLM 5
28.1
Qwen3.5 397B A17B
27.5
OpenCompass · IFEval
GLM 5 leads by +1.7
GLM 5
93.2
Qwen3.5 397B A17B
91.5
OpenCompass · LiveCodeBenchV6
GLM 5 leads by +3.2
GLM 5
86.2
Qwen3.5 397B A17B
83.0
OpenCompass · MMLU-Pro
Qwen3.5 397B A17B leads by +2.4
GLM 5
85.2
Qwen3.5 397B A17B
87.6
OTIS Mock AIME 2024-2025
Qwen3.5 397B A17B leads by +8.9
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
GLM 5
80.0
Qwen3.5 397B A17B
88.9
Full benchmark table
| Benchmark | GLM 5 | Qwen3.5 397B A17B |
|---|---|---|
APEX-Agents APEX-Agents · evaluates AI agents on complex, multi-step tasks requiring planning, tool use, and autonomous decision-making in realistic environments. | 17.2 | 24.9 |
Chatbot Arena Elo · Coding | 1434.0 | 1399.4 |
Chatbot Arena Elo · Overall | 1457.6 | 1441.8 |
Chess Puzzles Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities. | 5.3 | 8.5 |
GPQA diamond Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs. | 83.8 | 81.8 |
OpenCompass · AIME2025 | 95.8 | 92.3 |
OpenCompass · GPQA-Diamond | 85.3 | 88.4 |
OpenCompass · HLE | 28.1 | 27.5 |
OpenCompass · IFEval | 93.2 | 91.5 |
OpenCompass · LiveCodeBenchV6 | 86.2 | 83.0 |
OpenCompass · MMLU-Pro | 85.2 | 87.6 |
OTIS Mock AIME 2024-2025 OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills. | 80.0 | 88.9 |
Pricing · per 1M tokens · projected $/mo at 10M tokens
| Model | Input | Output | Context | Projected $/mo |
|---|---|---|---|---|
| $0.60 | $1.92 | 205K tokens (~102 books) | $9.30 | |
| $0.55 | $3.50 | 262K tokens (~131 books) | $12.88 |