Home/Models/Kimi K2 0711
moonshotai logo

Kimi K2 0711

by moonshotai · Released Jul 2025

Open Source
57.7
avg score
Rank #102
Compare
Better than 67% of all models
Context
131K tokens (~66 books)
Input $/1M
$0.57
Output $/1M
$2.30
Type
text
License
Open Source
Benchmarks
12 tested
Data as of
About

Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It is optimized for...

Tested on 12 benchmarks · BenchGecko score 57.7. Top scores: HELM — WildBench (86.2%), Lech Mazur Writing (85.6%), HELM — IFEval (85.0%).

Looking for similar performance at lower cost?
Kimi K2.5 scores 57.1 (99% as good) at $0.45/1M input · 21% cheaper
Capabilities
coding
32.8
#160 globally
reasoning
48.9
#81 globally
math
65.4
#61 globally
knowledge
73.5
#11 globally
language
85.0
#37 globally
Benchmark Scores
Compare All
Tested on 12 benchmarks · Ranked across 5 categories
Score Distribution (all 312 models)
0255075100
▲ You are here
Aider polyglot

Multi-language code editing from Aider. Tests editing ability across Python, JavaScript, TypeScript, Java, C++, Go, Rust, and more.

59.1·
WeirdML

Unusual and adversarial machine learning challenges. Tests robustness of reasoning about edge cases in ML systems.

39.4·
Terminal Bench

Complex terminal-based engineering tasks. Models must use command-line tools, navigate filesystems, and debug systems through shell interaction.

27.8·
HELM — WildBench

Stanford HELM WildBench evaluation. Tests reasoning on challenging real-world tasks.

86.2·
SimpleBench

Deceptively simple questions that humans find easy but AI models often get wrong. Tests common sense and reasoning gaps.

11.6·
HELM — Omni-MATH

Stanford HELM evaluation of mathematical reasoning across diverse problem types.

65.4·
Excellent (85+) Good (70-85) Average (50-70) Below (<50)
Links
Documentation
Community
BenchGecko API
kimi-k2
Specifications
  • Typetext
  • Context131K tokens (~66 books)
  • ReleasedJul 2025
  • LicenseOpen Source
  • StatusActive
  • Cost / Message~$0.003
Available On
moonshotai logomoonshotai$0.57
Share & Export
Tweet
Kimi K2 0711 is an open-source text AI model by moonshotai, released in July 2025. It has an average benchmark score of 57.7. Context window: 131K tokens.

Key facts · as of 2026-10-05

  • Kimi K2 0711 by moonshotai. BenchGecko score 57.7, rank 102 of 312 scored models (normalized average of public benchmark scores).
  • List price $0.57 input · $2.30 output per 1M tokens (as of 2026-10-05).
  • Sold by 1 provider (as of 2026-10-05): Novita (fp8) $0.57 in / $2.30 out.

How to cite · data as of 2026-10-05

Kimi K2 0711 · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/kimi-k2

Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP