Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...
Tested on 13 benchmarks · BenchGecko score 49.2. Top scores: MATH level 5 (97.1%), Lech Mazur Writing (87.1%), OTIS Mock AIME 2024-2025 (73.3%).
Phi 4 scores 49.0 (100% as good) at $0.07/1M input · 91% cheaper
Competition-level math from AMC, AIME, and olympiad problems. Level 5 is the hardest tier, requiring creative problem-solving.
Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.
Writing quality evaluation by Lech Mazur. Tests prose quality, coherence, and stylistic ability.
Simple factual questions with verified correct answers. Tests accuracy of basic knowledge retrieval. Low scores indicate hallucination.
LiveBench fiction analysis. Tests literary comprehension and creative text understanding.
- Typetext
- Context262K tokens (~131 books)
- ReleasedSep 2025
- LicenseOpen Source
- StatusActive
- Cost / Message~$0.005
Frequently Asked Questions
Key facts · as of 2026-10-05
- Qwen3 Max by Alibaba Qwen. BenchGecko score 49.2, rank 148 of 312 scored models (normalized average of public benchmark scores).
- List price $0.78 input · $3.90 output per 1M tokens (as of 2026-10-05).
- Sold by 1 provider (as of 2026-10-05): Alibaba $0.78 in / $3.90 out.
How to cite · data as of 2026-10-05
Qwen3 Max · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/qwen3-max
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP