Home/Models/Claude 3 Sonnet
Anthropic logo

Claude 3 Sonnet

by Anthropic · Released Jan 2024

30.3
avg score
Rank #246
Compare
Better than 21% of all models
Context
N/A
Input $/1M
n/a
Output $/1M
n/a
Type
text
License
Proprietary
Benchmarks
7 tested
Data as of
About

Tested on 7 benchmarks · BenchGecko score 30.3. Top scores: MMLU (67.9%), Winogrande (50.2%), Dtbench (22.7%).

Capabilities
coding
10.2
#189 globally
math
10.3
#254 globally
knowledge
46.3
#156 globally
general
22.7
#152 globally
Benchmark Scores
Compare All
Tested on 7 benchmarks · Ranked across 4 categories
Score Distribution (all 312 models)
0255075100
▲ You are here
WeirdML

Unusual and adversarial machine learning challenges. Tests robustness of reasoning about edge cases in ML systems.

10.2·
MATH level 5

Competition-level math from AMC, AIME, and olympiad problems. Level 5 is the hardest tier, requiring creative problem-solving.

18.2·
OTIS Mock AIME 2024-2025

Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.

2.4·
MMLU

Massive Multitask Language Understanding. 57 subjects from STEM, humanities, and social sciences. The most widely-cited knowledge benchmark.

67.9·
Winogrande

Commonsense coreference resolution. Tests understanding of pronoun references in ambiguous sentences.

50.2·
GPQA diamond

Graduate-level science questions written by PhD experts. Diamond subset contains questions where experts disagree, testing deep understanding.

20.8·
Excellent (85+) Good (70-85) Average (50-70) Below (<50)
Links
Documentation
Community
BenchGecko API
claude-3-sonnet
Specifications
  • Typetext
  • ContextN/A
  • ReleasedJan 2024
  • LicenseProprietary
  • Statusbenchmark-only
Available On
Anthropic logoAnthropicn/a
Share & Export
Tweet
Claude 3 Sonnet is a proprietary text AI model by Anthropic, released in January 2024. It has an average benchmark score of 30.3.

Key facts · as of 2026-04-09

  • Claude 3 Sonnet by Anthropic. BenchGecko score 30.3, rank 246 of 312 scored models (normalized average of public benchmark scores).
  • List price n/a input · n/a output per 1M tokens (as of 2026-04-09).

How to cite · data as of 2026-04-09

Claude 3 Sonnet · benchmarks, pricing and providers. BenchGecko, data as of 2026-04-09. https://benchgecko.ai/model/claude-3-sonnet

Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP