GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of GPT-4 Turbo while being twice as...
Tested on 8 benchmarks · BenchGecko score 58.5. Top scores: Chatbot Arena Elo — Overall (1346.1%), ScienceQA (84.7%), MMLU (78.9%).
DeepSeek V4.1 Flash scores 58.7 (100% as good) at $0.15/1M input · 97% cheaper
Code editing benchmark from the Aider project. Measures ability to apply targeted code changes while maintaining correctness and style.
Competition-level math from AMC, AIME, and olympiad problems. Level 5 is the hardest tier, requiring creative problem-solving.
Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.
Science questions with multimodal context including diagrams and charts from K-12 curriculum.
Massive Multitask Language Understanding. 57 subjects from STEM, humanities, and social sciences. The most widely-cited knowledge benchmark.
Broad Assessment of Language and Reasoning Over Games. Tests strategic and logical reasoning through game scenarios.
- Typemultimodal
- Context128K tokens (~64 books)
- ReleasedMay 2024
- LicenseProprietary
- StatusActive
- Cost / Message~$0.025
Frequently Asked Questions
Key facts · as of 2026-10-05
- GPT-4o (2024-05-13) by OpenAI. BenchGecko score 58.5, rank 97 of 312 scored models (normalized average of public benchmark scores).
- List price $5.00 input · $15.00 output per 1M tokens (as of 2026-10-05).
- Sold by 2 providers (as of 2026-10-05): Azure $5.00 in / $15.00 out · OpenAI $5.00 in / $15.00 out. Every provider
How to cite · data as of 2026-10-05
GPT-4o (2024-05-13) · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/gpt-4o-2024-05-13
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP