BenchmarksReading · ~3 min · 30 words deep

ARC-AGI

ARC-AGI is a reasoning benchmark tracked on BenchGecko. Abstraction and Reasoning Corpus. Tests fluid intelligence through novel visual pattern recognition puzzles. Core measure of general intelligence.

Built from BenchGecko data as of October 2, 2026 · updates when the data changes

ARC-AGI leaderboard
TL;DR

ARC-AGI is a reasoning benchmark tracked on BenchGecko. Abstraction and Reasoning Corpus. Tests fluid intelligence through novel visual pattern recognition puzzles. Core measure of general intelligence.

Level 1

Abstraction and Reasoning Corpus. Tests fluid intelligence through novel visual pattern recognition puzzles. Core measure of general intelligence.

Level 2

Scores are reported (%, maximum 100); higher is better. BenchGecko collects ARC-AGI scores from Epoch AI and refreshes them daily; the leaderboard on this page shows the current top models.

Level 3

Compare models only on the same benchmark version and settings. A benchmark separates models well while scores are spread out; when top models bunch together near the maximum, it stops telling them apart.

The takeaway for you
If you are a
Researcher
  • ·reasoning benchmark (%, maximum 100)
  • ·Source: Epoch AI
If you are a
Builder
  • ·Useful if your workload is reasoning
  • ·Check the live leaderboard on /benchmark/arc-agi before choosing a model
If you are a
Investor
  • ·Labs cite benchmark results in launch announcements; check the source and version
If you are a
Curious · Normie
  • ·ARC-AGI is a test of how good an AI is at reasoning
  • ·A higher score means better results on that kind of task
Abstraction and Reasoning Corpus. Tests fluid intelligence through novel visual pattern recognition puzzles. Core measure of general intelligence.