Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
Tested on 5 benchmarks. Top scores: Chatbot Arena Elo — Overall (1343.3%), Chatbot Arena Elo — Coding (1170.0%), Artificial Analysis — Agentic Index (39.7%).
Chatbot Arena overall Elo rating. Crowdsourced human preference ranking from blind head-to-head comparisons across all topics.
Chatbot Arena coding Elo. Human preference ranking specifically for coding tasks and technical questions.
Artificial Analysis Agentic Index. Composite score measuring agent capability across tool use and planning tasks.
Artificial Analysis Coding Index. Composite coding quality score from multiple code benchmarks.
Artificial Analysis Quality Index. Composite quality score combining multiple benchmark results into a single metric.
- Typetext
- Context128K tokens (~64 books)
- ReleasedMar 2026
- LicenseProprietary
- StatusActive
- Cost / Message~$0.001
Frequently Asked Questions
Related Models
Qwen3 Coder Next FP8 · AlibabaQwen2.5 Coder 32B Instruct AWQ · AlibabaSpeaker Diarization Community 1 · PyannoteGemma 3 12B (free) · Google DeepMindQwen3 Next 80B A3B Instruct (free) · Alibaba QwenKey facts · as of 2026-10-05
- Mercury 2 by inception. Not enough public benchmark scores to rank yet.
- List price $0.25 input · $0.75 output per 1M tokens (as of 2026-10-05).
- Sold by 1 provider (as of 2026-10-05): Inception $0.25 in / $0.75 out.
How to cite · data as of 2026-10-05
Mercury 2 · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/mercury-2
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP