Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
Tested on 6 benchmarks · BenchGecko score 39.2. Top scores: OTIS Mock AIME 2024-2025 (66.4%), Fiction.LiveBench (62.5%), GPQA diamond (51.7%).
ERNIE 4.5 21B A3B Thinking scores 39.8 (102% as good) at $0.07/1M input · 42% cheaper
Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.
LiveBench fiction analysis. Tests literary comprehension and creative text understanding.
Graduate-level science questions written by PhD experts. Diamond subset contains questions where experts disagree, testing deep understanding.
Tactical chess puzzles testing pattern recognition and multi-move calculation. Measures strategic reasoning ability.
- Typetext
- Context131K tokens (~66 books)
- ReleasedApr 2025
- LicenseOpen Source
- StatusActive
- Cost / Message~$0.000
Frequently Asked Questions
Key facts · as of 2026-10-05
- Qwen3 14B by Alibaba Qwen. BenchGecko score 39.2, rank 201 of 312 scored models (normalized average of public benchmark scores).
- List price $0.12 input · $0.24 output per 1M tokens (as of 2026-10-05).
- Sold by 3 providers (as of 2026-10-05): NextBit (int4) $0.10 in / $0.22 out · DeepInfra (fp8) $0.12 in / $0.24 out · Alibaba $0.23 in / $0.91 out. Every provider
How to cite · data as of 2026-10-05
Qwen3 14B · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/qwen3-14b
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP