Home/Models/GPT-5.4
OpenAI logo

GPT-5.4

by OpenAI · Released Mar 2026

Multimodal1M Context
66.0
avg score
Rank #54
Compare
Better than 83% of all models
Context
1.1M tokens (~525 books)
Input $/1M
$2.50
Output $/1M
$15.00
Type
multimodal
License
Proprietary
Benchmarks
34 tested
Data as of
About

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

Tested on 34 benchmarks · BenchGecko score 66.0. Top scores: Chatbot Arena Elo — Overall (1464.8%), Chatbot Arena Elo — Coding (1395.0%), OTIS Mock AIME 2024-2025 (97.8%).

Looking for similar performance at lower cost?
Gemini 3.7 Flash scores 66.4 (101% as good) at $0.75/1M input · 70% cheaper
Capabilities
coding
52.5
#79 globally
reasoning
83.8
#11 globally
math
60.0
#87 globally
knowledge
44.3
#171 globally
agentic
52.4
#18 globally
general
44.3
#52 globally
speed
61.3
#7 globally
Benchmark Scores
Compare All
Tested on 34 benchmarks · Ranked across 8 categories
Score Distribution (all 312 models)
0255075100
▲ You are here
Terminal Bench

Complex terminal-based engineering tasks. Models must use command-line tools, navigate filesystems, and debug systems through shell interaction.

81.8·
SWE-Bench verified

Real-world software engineering tasks from GitHub issues. Models must diagnose bugs and write patches that pass test suites. Human-verified subset of SWE-bench.

76.9·
WeirdML

Unusual and adversarial machine learning challenges. Tests robustness of reasoning about edge cases in ML systems.

57.4·
ARC-AGI

Abstraction and Reasoning Corpus. Tests fluid intelligence through novel visual pattern recognition puzzles. Core measure of general intelligence.

93.7·
ARC-AGI-2

ARC-AGI 2, harder sequel to ARC. More complex abstract reasoning patterns that test generalization ability beyond training data.

74.0·
OTIS Mock AIME 2024-2025

Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.

97.8·
Excellent (85+) Good (70-85) Average (50-70) Below (<50)
Links
Documentation
Community
BenchGecko API
gpt-5-4
Specifications
  • Typemultimodal
  • Context1.1M tokens (~525 books)
  • ReleasedMar 2026
  • LicenseProprietary
  • StatusActive
  • Cost / Message~$0.020
Available On
OpenAI logoOpenAI$2.50
Share & Export
Tweet
GPT-5.4 is a proprietary multimodal AI model by OpenAI, released in March 2026. It has an average benchmark score of 66.0. Context window: 1M tokens.

Key facts · as of 2026-10-05

  • GPT-5.4 by OpenAI. BenchGecko score 66.0, rank 54 of 312 scored models (normalized average of public benchmark scores).
  • List price $2.50 input · $15.00 output per 1M tokens (as of 2026-10-05).
  • Sold by 7 providers (as of 2026-10-05): OpenAI $1.25 in / $7.50 out · Azure $2.50 in / $15.00 out · OpenAI $2.50 in / $15.00 out · Amazon Bedrock $2.75 in / $16.50 out · Azure $2.75 in / $16.50 out · and 2 more. Every provider

How to cite · data as of 2026-10-05

GPT-5.4 · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/gpt-5-4

Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP