Mistral Large 2 2411 is an update of Mistral Large 2 released together with Pixtral Large 2411 It provides a significant upgrade on the previous Mistral Large 24.07, with notable...
Tested on 11 benchmarks · BenchGecko score 35.8. Top scores: Chatbot Arena Elo — Overall (1304.7%), HELM — IFEval (87.6%), HELM — WildBench (80.1%).
Mistral Nemo scores 35.4 (99% as good) at $0.02/1M input · 99% cheaper
Code editing benchmark from the Aider project. Measures ability to apply targeted code changes while maintaining correctness and style.
Stanford HELM WildBench evaluation. Tests reasoning on challenging real-world tasks.
Competition-level math from AMC, AIME, and olympiad problems. Level 5 is the hardest tier, requiring creative problem-solving.
Stanford HELM evaluation of mathematical reasoning across diverse problem types.
Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.
- Typetext
- Context131K tokens (~66 books)
- ReleasedNov 2024
- LicenseOpen Source
- StatusActive
- Cost / Message~$0.010
Frequently Asked Questions
Related Models
GPT-5.4 Mini · OpenAIClaude Haiku 4.5 · AnthropicLlama 3.2 3B Instruct · MetaMistral Nemo · Mistral AIQwen2.5-Max · Alibaba QwenKey facts · as of 2026-05-03
- Mistral Large 2411 by Mistral AI. BenchGecko score 35.8, rank 221 of 312 scored models (normalized average of public benchmark scores).
- List price $2.00 input · $6.00 output per 1M tokens (as of 2026-05-03).
How to cite · data as of 2026-05-03
Mistral Large 2411 · benchmarks, pricing and providers. BenchGecko, data as of 2026-05-03. https://benchgecko.ai/model/mistral-large-2411
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP