Home/Models/Qwen3 235B A22B Instruct 2507
Alibaba Qwen logo

Qwen3 235B A22B Instruct 2507

by Alibaba Qwen · Released Jul 2025

Open Source
44.9
avg score
Rank #172
Compare
Better than 45% of all models
Context
262K tokens (~131 books)
Input $/1M
$0.09
Output $/1M
$0.55
Type
text
License
Open Source
Benchmarks
22 tested
Data as of
About

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...

Tested on 22 benchmarks · BenchGecko score 44.9. Top scores: Chatbot Arena Elo — Overall (1422.4%), OpenCompass — IFEval (88.3%), OpenCompass — MMLU-Pro (79.2%).

Looking for similar performance at lower cost?
gpt-oss-20b scores 45.0 (100% as good) at $0.02/1M input · 80% cheaper
Capabilities
coding
44.8
#118 globally
reasoning
28.9
#123 globally
math
68.8
#51 globally
knowledge
53.7
#97 globally
general
45.8
#47 globally
language
58.7
#106 globally
Benchmark Scores
Compare All
Tested on 22 benchmarks · Ranked across 7 categories
Score Distribution (all 312 models)
0255075100
▲ You are here
LiveBench — Coding

Regularly refreshed coding problems that avoid data contamination. New problems added monthly to prevent memorization.

69.6·
Aider polyglot

Multi-language code editing from Aider. Tests editing ability across Python, JavaScript, TypeScript, Java, C++, Go, Rust, and more.

59.6·
OpenCompass — LiveCodeBenchV6

OpenCompass Live Code Bench v6. Fresh competitive programming problems to evaluate code generation without memorization.

43.0·
LiveBench — Reasoning

Regularly refreshed reasoning problems testing logical deduction, spatial reasoning, and analytical thinking.

58.4·
LiveBench — Data Analysis

Fresh data analysis tasks testing ability to interpret tables, charts, and statistical data.

44.7·
ARC-AGI

Abstraction and Reasoning Corpus. Tests fluid intelligence through novel visual pattern recognition puzzles. Core measure of general intelligence.

11.0·
OpenCompass — AIME2025

OpenCompass evaluation on AIME 2025 problems. Tests mathematical reasoning on fresh competition problems.

69.5·
LiveBench — Mathematics

Regularly updated math problems that test numerical reasoning, algebra, calculus, and combinatorics.

68.0·
Excellent (85+) Good (70-85) Average (50-70) Below (<50)
Model Family · Alibaba Qwen Qwen 3
See the full Qwen 3 family →
Recently Happened
Qwen3 235B A22B Instruct 2507 pricing increased 57%
Sep 5, 2026
Qwen3 235B A22B Instruct 2507 pricing dropped 36%
Sep 2, 2026
Qwen3 235B A22B Instruct 2507 pricing dropped 36%
Aug 27, 2026
Links
Documentation
Community
BenchGecko API
qwen3-235b-a22b-2507
Specifications
  • Typetext
  • Context262K tokens (~131 books)
  • ReleasedJul 2025
  • LicenseOpen Source
  • StatusActive
  • Cost / Message~$0.001
Available On
Alibaba Qwen logoAlibaba Qwen$0.09
Share & Export
Tweet
Qwen3 235B A22B Instruct 2507 is an open-source text AI model by Alibaba Qwen, released in July 2025. It has an average benchmark score of 44.9. Context window: 262K tokens.

Key facts · as of 2026-10-05

  • Qwen3 235B A22B Instruct 2507 by Alibaba Qwen. BenchGecko score 44.9, rank 172 of 312 scored models (normalized average of public benchmark scores).
  • List price $0.0900 input · $0.55 output per 1M tokens (as of 2026-10-05).
  • Sold by 9 providers (as of 2026-10-05): GMICloud (fp8) $0.0875 in / $0.35 out · DeepInfra (fp8) $0.0900 in / $0.55 out · Novita (fp8) $0.0900 in / $0.58 out · Parasail (fp8) $0.14 in / $0.80 out · Alibaba $0.15 in / $0.60 out · and 4 more. Every provider

How to cite · data as of 2026-10-05

Qwen3 235B A22B Instruct 2507 · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/qwen3-235b-a22b-2507

Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP