FrontierMath-Tiers-1-3-v2-Private
The Frontier
Best score over time · one chart, every benchmark
Full rankings
69 models tested · sorted by score
Score distribution
Where models cluster
Correlated benchmarks
Pearson r · original research
Benchmarks that track with FrontierMath-Tiers-1-3-v2-Private
Pearson correlation across models scored on both benchmarks. Closer to 1 = strongly predictive.
Frequently asked
About FrontierMath-Tiers-1-3-v2-Private
What does FrontierMath-Tiers-1-3-v2-Private measure?
FrontierMath-Tiers-1-3-v2-Private is a knowledge benchmark in the BenchGecko catalog. 69 AI models have been tested on it. Scores range from 0.3 to 93.7 out of 100.
Which model leads on FrontierMath-Tiers-1-3-v2-Private?
GPT-6 Astra from OpenAI leads FrontierMath-Tiers-1-3-v2-Private with a score of 93.7. The median score across 69 tested models is 55.4.
Is FrontierMath-Tiers-1-3-v2-Private saturated?
No · the top score is 93.7 out of 100 (94%). There is still meaningful room for improvement on FrontierMath-Tiers-1-3-v2-Private.
Does FrontierMath-Tiers-1-3-v2-Private predict performance on other benchmarks?
Yes · FrontierMath-Tiers-1-3-v2-Private scores correlate 0.98 with FrontierMath-2025-02-28-Private across 30 shared models. Models that do well on FrontierMath-Tiers-1-3-v2-Private tend to do well on FrontierMath-2025-02-28-Private.
How often is FrontierMath-Tiers-1-3-v2-Private data refreshed?
BenchGecko pulls updates daily. New model scores on FrontierMath-Tiers-1-3-v2-Private appear as soon as they are published by Epoch AI or the model provider.
- Category
- Knowledge
- Max score
- 100
- Models
- 69
- Updated
- 2026-09-28
Top on FrontierMath-Tiers-1-3-v2-Private
GPT-6 Astra · 93.7Claude Opus 5.5 · 91.2Claude Fable 5.1 · 90.2GPT-5.6 Sol · 89.1Claude Sonnet 5.5 · 88.8More knowledge benchmarks
Same category · related evaluations