Home/Models/Claude Sonnet 4.6
Anthropic logo

Claude Sonnet 4.6

by Anthropic · Released Feb 2026

Multimodal1M Context
52.5
avg score
Rank #130
Compare
Better than 58% of all models
Context
1.0M tokens (~500 books)
Input $/1M
$3.00
Output $/1M
$15.00
Type
multimodal
License
Proprietary
Benchmarks
30 tested
Data as of
About

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

Tested on 30 benchmarks · BenchGecko score 52.5. Top scores: Chatbot Arena Elo — Coding (1521.4%), Chatbot Arena Elo — Overall (1472.2%), ARC-AGI (86.5%).

Looking for similar performance at lower cost?
Gemma 4 31B scores 53.3 (102% as good) at $0.09/1M input · 97% cheaper
Capabilities
coding
49.8
#93 globally
reasoning
73.5
#35 globally
math
44.0
#141 globally
knowledge
40.6
#201 globally
agentic
44.5
#25 globally
general
37.2
#82 globally
speed
50.3
#30 globally
Benchmark Scores
Compare All
Tested on 30 benchmarks · Ranked across 8 categories
Score Distribution (all 312 models)
0255075100
▲ You are here
SWE-Bench verified

Real-world software engineering tasks from GitHub issues. Models must diagnose bugs and write patches that pass test suites. Human-verified subset of SWE-bench.

75.2·
WeirdML

Unusual and adversarial machine learning challenges. Tests robustness of reasoning about edge cases in ML systems.

66.1·
Terminal Bench

Complex terminal-based engineering tasks. Models must use command-line tools, navigate filesystems, and debug systems through shell interaction.

53.4·
ARC-AGI

Abstraction and Reasoning Corpus. Tests fluid intelligence through novel visual pattern recognition puzzles. Core measure of general intelligence.

86.5·
ARC-AGI-2

ARC-AGI 2, harder sequel to ARC. More complex abstract reasoning patterns that test generalization ability beyond training data.

60.4·
OTIS Mock AIME 2024-2025

Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.

85.8·
FrontierMath-2025-02-28-Private

Original research-level math problems created by professional mathematicians. Problems are unpublished and cannot be memorized.

32.4·
FrontierMath-Tier-4-2025-07-01-Private

Hardest tier of FrontierMath. Problems at the frontier of human mathematical ability, many unsolved by most mathematicians.

13.8·
Excellent (85+) Good (70-85) Average (50-70) Below (<50)
Links
Documentation
Community
BenchGecko API
claude-sonnet-4-6
Specifications
  • Typemultimodal
  • Context1.0M tokens (~500 books)
  • ReleasedFeb 2026
  • LicenseProprietary
  • StatusActive
  • Cost / Message~$0.021
Available On
Anthropic logoAnthropic$3.00
Share & Export
Tweet
Claude Sonnet 4.6 is a proprietary multimodal AI model by Anthropic, released in February 2026. It has an average benchmark score of 52.5. Context window: 1M tokens.

Key facts · as of 2026-10-05

  • Claude Sonnet 4.6 by Anthropic. BenchGecko score 52.5, rank 130 of 312 scored models (normalized average of public benchmark scores).
  • List price $3.00 input · $15.00 output per 1M tokens (as of 2026-10-05).
  • Sold by 9 providers (as of 2026-10-05): Amazon Bedrock $3.00 in / $15.00 out · Anthropic $3.00 in / $15.00 out · Azure $3.00 in / $15.00 out · Claude Platform on AWS $3.00 in / $15.00 out · Google $3.00 in / $15.00 out · and 4 more. Every provider

How to cite · data as of 2026-10-05

Claude Sonnet 4.6 · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/claude-sonnet-4-6

Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP