GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex
Tested on 8 benchmarks · BenchGecko score 63.3. Top scores: LiveBench — Mathematics (76.3%), LiveBench — Coding (69.9%), LiveBench — Reasoning (64.7%).
Qwen3 30B A3B Thinking 2507 scores 63.5 (100% as good) at $0.20/1M input · 20% cheaper
Regularly refreshed coding problems that avoid data contamination. New problems added monthly to prevent memorization.
LiveBench coding tasks that require multi-step reasoning and tool use. Tests planning and execution of complex coding workflows.
Regularly refreshed reasoning problems testing logical deduction, spatial reasoning, and analytical thinking.
Fresh data analysis tasks testing ability to interpret tables, charts, and statistical data.
Regularly updated math problems that test numerical reasoning, algebra, calculus, and combinatorics.
- Typemultimodal
- Context400K tokens (~200 books)
- ReleasedNov 2025
- LicenseProprietary
- StatusActive
- Cost / Message~$0.003
Frequently Asked Questions
Related Models
GLM 5.3 · z-aiQwen3 30B A3B Thinking 2507 · Alibaba QwenKimi K2.7 Code · moonshotaiMixtral 8x7B · Mistral AIQwen3.6 Plus · Alibaba QwenKey facts · as of 2026-10-05
- GPT-5.1-Codex-Mini by OpenAI. BenchGecko score 63.3, rank 70 of 312 scored models (normalized average of public benchmark scores).
- List price $0.25 input · $2.00 output per 1M tokens (as of 2026-10-05).
- Sold by 1 provider (as of 2026-10-05): Azure $0.25 in / $2.00 out.
How to cite · data as of 2026-10-05
GPT-5.1-Codex-Mini · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/gpt-5-1-codex-mini
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP