QwQ is the reasoning model of the Qwen series. Compared with conventional instruction-tuned models, QwQ, which is capable of thinking and reasoning, can achieve significantly enhanced performance in downstream tasks,...
Tested on 13 benchmarks · BenchGecko score 33.7. Top scores: Chatbot Arena Elo — Overall (1335.8%), Fiction.LiveBench (83.3%), Lech Mazur Writing (80.2%).
Qwen3 30B A3B scores 34.0 (101% as good) at $0.12/1M input · 20% cheaper
Multi-language code editing from Aider. Tests editing ability across Python, JavaScript, TypeScript, Java, C++, Go, Rust, and more.
HuggingFace MuSR (Multi-Step Reasoning). Tests multi-hop reasoning requiring chaining multiple facts together.
Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.
HuggingFace evaluation of MATH Level 5 problems. Competition math requiring advanced reasoning and proof construction.
- Typetext
- Context131K tokens (~66 books)
- ReleasedMar 2025
- LicenseOpen Source
- StatusActive
- Cost / Message~$0.001
Frequently Asked Questions
Related Models
Qwen3 30B A3B · Alibaba QwenGPT-4o-mini (2024-07-18) · OpenAILLaMA-13B · MetaGemma 2B · Google DeepMindGemini 2.0 Flash Thinking (Jan 2025) · Google DeepMindKey facts · as of 2026-04-28
- QwQ 32B by Alibaba Qwen. BenchGecko score 33.7, rank 230 of 312 scored models (normalized average of public benchmark scores).
- List price $0.15 input · $0.58 output per 1M tokens (as of 2026-04-28).
How to cite · data as of 2026-04-28
QwQ 32B · benchmarks, pricing and providers. BenchGecko, data as of 2026-04-28. https://benchgecko.ai/model/qwq-32b
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP