Home/Models/Qwen3 4B Thinking 2507
Alibaba logo

Qwen3 4B Thinking 2507

by Alibaba · Released Aug 2025

Open Source
48.4
avg score
Rank #156
Compare
Better than 50% of all models
Context
N/A
Input $/1M
n/a
Output $/1M
n/a
Type
text-generation
License
Open Source
Benchmarks
6 tested
Data as of
About

Qwen text generation model. 1217K downloads on HuggingFace.

Tested on 6 benchmarks · BenchGecko score 48.4. Top scores: OpenCompass — IFEval (88.5%), OpenCompass — AIME2025 (80.0%), OpenCompass — MMLU-Pro (72.8%).

Capabilities
coding
51.6
#87 globally
math
80.0
#25 globally
knowledge
47.8
#144 globally
language
88.5
#19 globally
Benchmark Scores
Compare All
Tested on 6 benchmarks · Ranked across 4 categories
Score Distribution (all 312 models)
0255075100
▲ You are here
OpenCompass — LiveCodeBenchV6

OpenCompass Live Code Bench v6. Fresh competitive programming problems to evaluate code generation without memorization.

51.6·
OpenCompass — AIME2025

OpenCompass evaluation on AIME 2025 problems. Tests mathematical reasoning on fresh competition problems.

80.0·
OpenCompass — MMLU-Pro

OpenCompass MMLU-Pro evaluation. Harder knowledge test with more answer choices.

72.8·
OpenCompass — GPQA-Diamond

OpenCompass evaluation of GPQA Diamond. PhD-level science questions from the hardest subset.

64.7·
OpenCompass — HLE

OpenCompass evaluation of Humanitys Last Exam. Expert-level cross-discipline knowledge test.

6.0·
Excellent (85+) Good (70-85) Average (50-70) Below (<50)
Links
Documentation
Community
BenchGecko API
qwen-qwen3-4b-thinking-2507
Specifications
  • Typetext-generation
  • ContextN/A
  • ReleasedAug 2025
  • LicenseOpen Source
  • StatusActive
Available On
Alibaba logoAlibaban/a
Share & Export
Tweet
Qwen3 4B Thinking 2507 is an open-source text-generation AI model by Alibaba, released in August 2025. It has an average benchmark score of 48.4.

Key facts · as of 2026-04-09

  • Qwen3 4B Thinking 2507 by Alibaba. BenchGecko score 48.4, rank 154 of 312 scored models (normalized average of public benchmark scores).
  • List price n/a input · n/a output per 1M tokens (as of 2026-04-09).

How to cite · data as of 2026-04-09

Qwen3 4B Thinking 2507 · benchmarks, pricing and providers. BenchGecko, data as of 2026-04-09. https://benchgecko.ai/model/qwen-qwen3-4b-thinking-2507

Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP