Grok 3 Mini is a lightweight, smaller thinking model. Unlike traditional models that generate answers immediately, Grok 3 Mini thinks before responding. It’s ideal for reasoning-heavy tasks that don’t demand...
Tested on 7 benchmarks · BenchGecko score 53.1. Top scores: Chatbot Arena Elo — Overall (1357.4%), HELM — IFEval (95.1%), HELM — MMLU-Pro (79.9%).
Gemma 4 31B scores 53.3 (100% as good) at $0.09/1M input · 70% cheaper
Multi-language code editing from Aider. Tests editing ability across Python, JavaScript, TypeScript, Java, C++, Go, Rust, and more.
Stanford HELM WildBench evaluation. Tests reasoning on challenging real-world tasks.
Stanford HELM evaluation of mathematical reasoning across diverse problem types.
- Typetext
- Context131K tokens (~66 books)
- ReleasedApr 2025
- LicenseProprietary
- Statuspreview
- Cost / Message~$0.001
Frequently Asked Questions
Key facts · as of 2026-05-03
- Grok 3 Mini Beta by xAI. BenchGecko score 53.1, rank 128 of 312 scored models (normalized average of public benchmark scores).
- List price $0.30 input · $0.50 output per 1M tokens (as of 2026-05-03).
How to cite · data as of 2026-05-03
Grok 3 Mini Beta · benchmarks, pricing and providers. BenchGecko, data as of 2026-05-03. https://benchgecko.ai/model/grok-3-mini-beta
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP