WizardLM-2 8x22B is Microsoft AI's most advanced Wizard model. It demonstrates highly competitive performance compared to leading proprietary models, and it consistently outperforms all existing state-of-the-art opensource models. It is...
Tested on 7 benchmarks · BenchGecko score 61.4. Top scores: IFEval (52.7%), BBH (HuggingFace) (48.6%), Aider — Code Editing (44.4%).
gpt-oss-20b (free) scores 61.0 (99% as good) at $0.00/1M input · 100% cheaper
Code editing benchmark from the Aider project. Measures ability to apply targeted code changes while maintaining correctness and style.
HuggingFace MuSR (Multi-Step Reasoning). Tests multi-hop reasoning requiring chaining multiple facts together.
HuggingFace evaluation of MATH Level 5 problems. Competition math requiring advanced reasoning and proof construction.
- Typetext
- Context66K tokens (~33 books)
- ReleasedApr 2024
- LicenseOpen Source
- StatusActive
- Cost / Message~$0.002
Frequently Asked Questions
Key facts · as of 2026-10-05
- WizardLM-2 8x22B by Microsoft. BenchGecko score 61.4, rank 78 of 312 scored models (normalized average of public benchmark scores).
- List price $0.62 input · $0.62 output per 1M tokens (as of 2026-10-05).
- Sold by 1 provider (as of 2026-10-05): Novita (bf16) $0.62 in / $0.62 out.
How to cite · data as of 2026-10-05
WizardLM-2 8x22B · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/wizardlm-2-8x22b
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP