Tested on 5 benchmarks · BenchGecko score 68.7. Top scores: GSM8K (86.7%), ARC AI2 (81.7%), TriviaQA (78.9%).
Grade school math word problems. 8,500 problems testing multi-step arithmetic reasoning. A foundational math benchmark.
AI2 Reasoning Challenge. Grade-school science questions requiring multi-step reasoning. Easy and Challenge sets test different difficulty levels.
Trivia questions sourced from trivia enthusiasts and quiz websites. Tests breadth of general knowledge.
Massive Multitask Language Understanding. 57 subjects from STEM, humanities, and social sciences. The most widely-cited knowledge benchmark.
- Typetext
- ContextN/A
- ReleasedJan 2024
- LicenseProprietary
- Statusbenchmark-only
Frequently Asked Questions
Key facts · as of 2026-04-09
- Claude Instant by Anthropic. BenchGecko score 68.7, rank 42 of 312 scored models (normalized average of public benchmark scores).
- List price n/a input · n/a output per 1M tokens (as of 2026-04-09).
How to cite · data as of 2026-04-09
Claude Instant · benchmarks, pricing and providers. BenchGecko, data as of 2026-04-09. https://benchgecko.ai/model/claude-instant
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP