GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
Tested on 22 benchmarks · BenchGecko score 35.2. Top scores: HELM — IFEval (90.4%), MATH level 5 (87.3%), HELM — WildBench (83.8%).
Gemma 3 27B (free) scores 35.1 (100% as good) at $0.00/1M input · 100% cheaper
Unusual and adversarial machine learning challenges. Tests robustness of reasoning about edge cases in ML systems.
Multi-language code editing from Aider. Tests editing ability across Python, JavaScript, TypeScript, Java, C++, Go, Rust, and more.
SWE-bench Verified solved using only bash commands, no specialized frameworks. Tests raw terminal-based problem solving.
Stanford HELM WildBench evaluation. Tests reasoning on challenging real-world tasks.
Abstraction and Reasoning Corpus. Tests fluid intelligence through novel visual pattern recognition puzzles. Core measure of general intelligence.
ARC-AGI 2, harder sequel to ARC. More complex abstract reasoning patterns that test generalization ability beyond training data.
Competition-level math from AMC, AIME, and olympiad problems. Level 5 is the hardest tier, requiring creative problem-solving.
Stanford HELM evaluation of mathematical reasoning across diverse problem types.
Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.
- Typemultimodal
- Context1.0M tokens (~524 books)
- ReleasedApr 2025
- LicenseProprietary
- StatusActive
- Cost / Message~$0.002
Frequently Asked Questions
Key facts · as of 2026-10-05
- GPT-4.1 Mini by OpenAI. BenchGecko score 35.2, rank 224 of 312 scored models (normalized average of public benchmark scores).
- List price $0.40 input · $1.60 output per 1M tokens (as of 2026-10-05).
- Sold by 3 providers (as of 2026-10-05): Azure $0.40 in / $1.60 out · OpenAI $0.40 in / $1.60 out · Azure $0.44 in / $1.76 out. Every provider
How to cite · data as of 2026-10-05
GPT-4.1 Mini · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/gpt-4-1-mini
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP