Compare · ModelsLive · 2 picked · head to head

GLM 5 vs Qwen3.5 397B A17B

Side by side · benchmarks, pricing, and signals you can act on.

Winner summary

GLM 5 wins 7 of 12 shared benchmarks. Leads in arena · math · language.

Category leads
agentic·Qwen3.5 397B A17Barena·GLM 5knowledge·Qwen3.5 397B A17Bmath·GLM 5language·GLM 5coding·GLM 5
Hype vs Reality
GLM 5
#79 by perf·#3 by attention
DESERVED
Qwen3.5 397B A17B
#48 by perf·#2 by attention
DESERVED
Best value
1.5x better value than Qwen3.5 397B A17B
GLM 5
43.7 pts/$
$1.26/M
Qwen3.5 397B A17B
29.6 pts/$
$2.02/M
Vendor risk
z-ai logo
z-ai
private · undisclosed
Unknown
Alibaba Qwen logo
Alibaba (Qwen)
$293.0B·Tier 1
Low risk
Head to head
GLM 5Qwen3.5 397B A17B
APEX-Agents
Qwen3.5 397B A17B leads by +7.7
APEX-Agents · evaluates AI agents on complex, multi-step tasks requiring planning, tool use, and autonomous decision-making in realistic environments.
GLM 5
17.2
Qwen3.5 397B A17B
24.9
Chatbot Arena Elo · Coding
GLM 5 leads by +34.6
GLM 5
1434.0
Qwen3.5 397B A17B
1399.4
Chatbot Arena Elo · Overall
GLM 5 leads by +15.8
GLM 5
1457.6
Qwen3.5 397B A17B
1441.8
Chess Puzzles
Qwen3.5 397B A17B leads by +3.2
Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities.
GLM 5
5.3
Qwen3.5 397B A17B
8.5
GPQA diamond
GLM 5 leads by +1.9
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
GLM 5
83.8
Qwen3.5 397B A17B
81.8
OpenCompass · AIME2025
GLM 5 leads by +3.5
GLM 5
95.8
Qwen3.5 397B A17B
92.3
OpenCompass · GPQA-Diamond
Qwen3.5 397B A17B leads by +3.1
GLM 5
85.3
Qwen3.5 397B A17B
88.4
OpenCompass · HLE
GLM 5 leads by +0.6
GLM 5
28.1
Qwen3.5 397B A17B
27.5
OpenCompass · IFEval
GLM 5 leads by +1.7
GLM 5
93.2
Qwen3.5 397B A17B
91.5
OpenCompass · LiveCodeBenchV6
GLM 5 leads by +3.2
GLM 5
86.2
Qwen3.5 397B A17B
83.0
OpenCompass · MMLU-Pro
Qwen3.5 397B A17B leads by +2.4
GLM 5
85.2
Qwen3.5 397B A17B
87.6
OTIS Mock AIME 2024-2025
Qwen3.5 397B A17B leads by +8.9
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
GLM 5
80.0
Qwen3.5 397B A17B
88.9
Full benchmark table
BenchmarkGLM 5Qwen3.5 397B A17B
APEX-Agents
APEX-Agents · evaluates AI agents on complex, multi-step tasks requiring planning, tool use, and autonomous decision-making in realistic environments.
17.224.9
Chatbot Arena Elo · Coding
1434.01399.4
Chatbot Arena Elo · Overall
1457.61441.8
Chess Puzzles
Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities.
5.38.5
GPQA diamond
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
83.881.8
OpenCompass · AIME2025
95.892.3
OpenCompass · GPQA-Diamond
85.388.4
OpenCompass · HLE
28.127.5
OpenCompass · IFEval
93.291.5
OpenCompass · LiveCodeBenchV6
86.283.0
OpenCompass · MMLU-Pro
85.287.6
OTIS Mock AIME 2024-2025
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
80.088.9
Pricing · per 1M tokens · projected $/mo at 10M tokens
ModelInputOutputContextProjected $/mo
z-ai logoGLM 5$0.60$1.92205K tokens (~102 books)$9.30
Alibaba Qwen logoQwen3.5 397B A17B$0.55$3.50262K tokens (~131 books)$12.88