Home/Models/Qwen3 235B A22B Thinking 2507
Alibaba Qwen logo

Qwen3 235B A22B Thinking 2507

by Alibaba Qwen · Released Jul 2025

Open Source
55.3
avg score
Rank #114
Compare
Better than 63% of all models
Context
131K tokens (~66 books)
Input $/1M
$0.23
Output $/1M
$2.30
Type
text
License
Open Source
Benchmarks
27 tested
Data as of
About

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...

Tested on 27 benchmarks · BenchGecko score 55.3. Top scores: Chatbot Arena Elo — Overall (1400.0%), OpenCompass — AIME2025 (90.9%), OpenCompass — IFEval (87.8%).

Capabilities
coding
46.8
#107 globally
reasoning
55.8
#69 globally
math
53.2
#109 globally
knowledge
57.3
#77 globally
general
33.9
#111 globally
language
66.0
#94 globally
Benchmark Scores
Compare All
Tested on 27 benchmarks · Ranked across 7 categories
Score Distribution (all 312 models)
0255075100
▲ You are here
OpenCompass — LiveCodeBenchV6

OpenCompass Live Code Bench v6. Fresh competitive programming problems to evaluate code generation without memorization.

70.6·
LiveBench — Coding

Regularly refreshed coding problems that avoid data contamination. New problems added monthly to prevent memorization.

69.0·
WeirdML

Unusual and adversarial machine learning challenges. Tests robustness of reasoning about edge cases in ML systems.

41.0·
LiveBench — Reasoning

Regularly refreshed reasoning problems testing logical deduction, spatial reasoning, and analytical thinking.

59.4·
LiveBench — Data Analysis

Fresh data analysis tasks testing ability to interpret tables, charts, and statistical data.

52.2·
OpenCompass — AIME2025

OpenCompass evaluation on AIME 2025 problems. Tests mathematical reasoning on fresh competition problems.

90.9·
OTIS Mock AIME 2024-2025

Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.

86.7·
LiveBench — Mathematics

Regularly updated math problems that test numerical reasoning, algebra, calculus, and combinatorics.

73.4·
Excellent (85+) Good (70-85) Average (50-70) Below (<50)
Model Family · Alibaba Qwen Qwen 3
See the full Qwen 3 family →
Recently Happened
Qwen3 235B A22B Thinking 2507 pricing increased 1395%
Jul 1, 2026
Qwen3 235B A22B Thinking 2507 pricing increased 149%
Apr 23, 2026
Qwen3 235B A22B Thinking 2507 pricing dropped 60%
Apr 17, 2026
Links
Documentation
Community
BenchGecko API
qwen3-235b-a22b-thinking-2507
Specifications
  • Typetext
  • Context131K tokens (~66 books)
  • ReleasedJul 2025
  • LicenseOpen Source
  • StatusActive
  • Cost / Message~$0.003
Available On
Alibaba Qwen logoAlibaba Qwen$0.23
Share & Export
Tweet
Qwen3 235B A22B Thinking 2507 is an open-source text AI model by Alibaba Qwen, released in July 2025. It has an average benchmark score of 55.3. Context window: 131K tokens.

Key facts · as of 2026-10-05

  • Qwen3 235B A22B Thinking 2507 by Alibaba Qwen. BenchGecko score 55.3, rank 114 of 312 scored models (normalized average of public benchmark scores).
  • List price $0.23 input · $2.30 output per 1M tokens (as of 2026-10-05).
  • Sold by 3 providers (as of 2026-10-05): Alibaba $0.23 in / $2.30 out · Novita (fp8) $0.30 in / $3.00 out · Venice (fp8) $0.45 in / $3.50 out. Every provider

How to cite · data as of 2026-10-05

Qwen3 235B A22B Thinking 2507 · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/qwen3-235b-a22b-thinking-2507

Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP