Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...
Tested on 22 benchmarks · BenchGecko score 44.9. Top scores: Chatbot Arena Elo — Overall (1422.4%), OpenCompass — IFEval (88.3%), OpenCompass — MMLU-Pro (79.2%).
gpt-oss-20b scores 45.0 (100% as good) at $0.02/1M input · 80% cheaper
Regularly refreshed coding problems that avoid data contamination. New problems added monthly to prevent memorization.
Multi-language code editing from Aider. Tests editing ability across Python, JavaScript, TypeScript, Java, C++, Go, Rust, and more.
OpenCompass Live Code Bench v6. Fresh competitive programming problems to evaluate code generation without memorization.
Regularly refreshed reasoning problems testing logical deduction, spatial reasoning, and analytical thinking.
Fresh data analysis tasks testing ability to interpret tables, charts, and statistical data.
Abstraction and Reasoning Corpus. Tests fluid intelligence through novel visual pattern recognition puzzles. Core measure of general intelligence.
OpenCompass evaluation on AIME 2025 problems. Tests mathematical reasoning on fresh competition problems.
Regularly updated math problems that test numerical reasoning, algebra, calculus, and combinatorics.
- Typetext
- Context262K tokens (~131 books)
- ReleasedJul 2025
- LicenseOpen Source
- StatusActive
- Cost / Message~$0.001
Frequently Asked Questions
Related Models
gpt-oss-20b · OpenAIQwen3.7 Flash · Alibaba QwenClaude Sonnet 4.5 · AnthropicGrok 3 · xAIClaude Sonnet 4 · AnthropicKey facts · as of 2026-10-05
- Qwen3 235B A22B Instruct 2507 by Alibaba Qwen. BenchGecko score 44.9, rank 172 of 312 scored models (normalized average of public benchmark scores).
- List price $0.0900 input · $0.55 output per 1M tokens (as of 2026-10-05).
- Sold by 9 providers (as of 2026-10-05): GMICloud (fp8) $0.0875 in / $0.35 out · DeepInfra (fp8) $0.0900 in / $0.55 out · Novita (fp8) $0.0900 in / $0.58 out · Parasail (fp8) $0.14 in / $0.80 out · Alibaba $0.15 in / $0.60 out · and 4 more. Every provider
How to cite · data as of 2026-10-05
Qwen3 235B A22B Instruct 2507 · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/qwen3-235b-a22b-2507
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP