SimpleBench
SimpleBench is a reasoning benchmark tracked on BenchGecko. Deceptively simple questions that humans find easy but AI models often get wrong. Tests common sense and reasoning gaps.
Built from BenchGecko data as of October 2, 2026 · updates when the data changes
SimpleBench is a reasoning benchmark tracked on BenchGecko. Deceptively simple questions that humans find easy but AI models often get wrong. Tests common sense and reasoning gaps.
Basic
Deceptively simple questions that humans find easy but AI models often get wrong. Tests common sense and reasoning gaps.
Deep
Scores are reported (%, maximum 100); higher is better. BenchGecko collects SimpleBench scores from Epoch AI and refreshes them daily; the leaderboard on this page shows the current top models.
Expert
Compare models only on the same benchmark version and settings. A benchmark separates models well while scores are spread out; when top models bunch together near the maximum, it stops telling them apart.
Depending on why you're here
- ·reasoning benchmark (%, maximum 100)
- ·Source: Epoch AI
- ·Useful if your workload is reasoning
- ·Check the live leaderboard on /benchmark/simplebench before choosing a model
- ·Labs cite benchmark results in launch announcements; check the source and version
- ·SimpleBench is a test of how good an AI is at reasoning
- ·A higher score means better results on that kind of task