Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...
Tested on 24 benchmarks · BenchGecko score 96.8. Top scores: OTIS Mock AIME 2024-2025 (100.0%), Proofbench (100.0%), Dtbench (98.2%).
Step 3.5 Flash scores 89.5 (92% as good) at $0.10/1M input · 98% cheaper
Abstraction and Reasoning Corpus. Tests fluid intelligence through novel visual pattern recognition puzzles. Core measure of general intelligence.
ARC-AGI 2, harder sequel to ARC. More complex abstract reasoning patterns that test generalization ability beyond training data.
Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.
- Typemultimodal
- Context1.0M tokens (~500 books)
- ReleasedSep 2026
- LicenseProprietary
- StatusActive
- Cost / Message~$0.028
Frequently Asked Questions
Key facts · as of 2026-10-05
- Claude Opus 5.5 by Anthropic. BenchGecko score 96.8, rank 3 of 312 scored models (normalized average of public benchmark scores).
- List price $4.00 input · $20.00 output per 1M tokens (as of 2026-10-05).
- Sold by 11 providers (as of 2026-10-05): Amazon Bedrock $4.00 in / $20.00 out · Anthropic $4.00 in / $20.00 out · Azure $4.00 in / $20.00 out · Claude Platform on AWS $4.00 in / $20.00 out · Google $4.00 in / $20.00 out · and 6 more. Every provider
- Gecko Tests · Gecko Score 65 (rank 3): Who Are You B (Knows who made it) · Censorship Index A (Answers almost everything) · Knowledge Horizon A (Knows news up to May 2026) · Tokenizer Tax E (73% more tokens outside English).
How to cite · data as of 2026-10-05
Claude Opus 5.5 · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/claude-opus-5-5
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP