Tested on 7 benchmarks · BenchGecko score 53.1. Top scores: Fiction.LiveBench (94.4%), Lech Mazur Writing (81.1%), Dtbench (71.1%).
Unusual and adversarial machine learning challenges. Tests robustness of reasoning about edge cases in ML systems.
Capture-the-flag cybersecurity challenges. Tests vulnerability analysis, reverse engineering, cryptography, and exploitation skills.
Abstraction and Reasoning Corpus. Tests fluid intelligence through novel visual pattern recognition puzzles. Core measure of general intelligence.
ARC-AGI 2, harder sequel to ARC. More complex abstract reasoning patterns that test generalization ability beyond training data.
LiveBench fiction analysis. Tests literary comprehension and creative text understanding.
Writing quality evaluation by Lech Mazur. Tests prose quality, coherence, and stylistic ability.
- Typetext
- ContextN/A
- ReleasedJan 2024
- LicenseProprietary
- Statusbenchmark-only
Frequently Asked Questions
Key facts · as of 2026-05-03
- Grok 4 Fast by xAI. BenchGecko score 53.1, rank 129 of 312 scored models (normalized average of public benchmark scores).
- List price n/a input · n/a output per 1M tokens (as of 2026-05-03).
How to cite · data as of 2026-05-03
Grok 4 Fast · benchmarks, pricing and providers. BenchGecko, data as of 2026-05-03. https://benchgecko.ai/model/grok-4-fast
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP