Qwen3-4B is a 4 billion parameter dense language model from the Qwen3 series, designed to support both general-purpose and reasoning-intensive tasks. It introduces a dual-mode architecture—thinking and non-thinking—allowing dynamic switching between high-precision logical reasoning and efficient dialogue generation. This makes it well-suited for multi-turn chat, instruction following, and complex agent workflows.
No benchmark data available yet.
- Typetext
- Context41K tokens (~20 books)
- ReleasedApr 2025
- LicenseOpen Source
- StatusActive
- Cost / Message~$0.000
Frequently Asked Questions
Related Models
Qwen3 Coder Next FP8 · AlibabaQwen2.5 Coder 32B Instruct AWQ · AlibabaSpeaker Diarization Community 1 · PyannoteGemma 3 12B (free) · Google DeepMindQwen3 Next 80B A3B Instruct (free) · Alibaba QwenBenchmarks
Key facts · as of 2026-03-27
- Qwen3 4B (free) by Alibaba Qwen. Not enough public benchmark scores to rank yet.
- List price free input · free output per 1M tokens (as of 2026-03-27).
How to cite · data as of 2026-03-27
Qwen3 4B (free) · benchmarks, pricing and providers. BenchGecko, data as of 2026-03-27. https://benchgecko.ai/model/qwen3-4b-free
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP