BenchmarksReading · ~3 min · 30 words deep

FrontierMath-Tier-4-2025-07-01-Private

FrontierMath-Tier-4-2025-07-01-Private is a math benchmark tracked on BenchGecko. Hardest tier of FrontierMath. Problems at the frontier of human mathematical ability, many unsolved by most mathematicians.

Built from BenchGecko data as of June 22, 2026 · updates when the data changes

TL;DR

FrontierMath-Tier-4-2025-07-01-Private is a math benchmark tracked on BenchGecko. Hardest tier of FrontierMath. Problems at the frontier of human mathematical ability, many unsolved by most mathematicians.

Live data · updated daily
See full leaderboard
Level 1

Hardest tier of FrontierMath. Problems at the frontier of human mathematical ability, many unsolved by most mathematicians.

Level 2

Scores are reported (%, maximum 100); higher is better. BenchGecko collects FrontierMath-Tier-4-2025-07-01-Private scores from Epoch AI and refreshes them daily; the leaderboard on this page shows the current top models.

Level 3

Compare models only on the same benchmark version and settings. A benchmark separates models well while scores are spread out; when top models bunch together near the maximum, it stops telling them apart.

The takeaway for you
If you are a
Researcher
  • ·math benchmark (%, maximum 100)
  • ·Source: Epoch AI
If you are a
Builder
  • ·Useful if your workload is math
  • ·Check the live leaderboard on /benchmark/frontiermath-tier-4-2025-07-01-private before choosing a model
If you are a
Investor
  • ·Labs cite benchmark results in launch announcements; check the source and version
If you are a
Curious · Normie
  • ·FrontierMath-Tier-4-2025-07-01-Private is a test of how good an AI is at math
  • ·A higher score means better results on that kind of task
Hardest tier of FrontierMath. Problems at the frontier of human mathematical ability, many unsolved by most mathematicians.