Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
Tested on 23 benchmarks · BenchGecko score 43.7. Top scores: Chatbot Arena Elo — Coding (1508.7%), Chatbot Arena Elo — Overall (1461.0%), OTIS Mock AIME 2024-2025 (96.1%).
Gemini 2.0 Flash scores 44.4 (102% as good) at $0.10/1M input · 89% cheaper
Real-world software engineering tasks from GitHub issues. Models must diagnose bugs and write patches that pass test suites. Human-verified subset of SWE-bench.
Unusual and adversarial machine learning challenges. Tests robustness of reasoning about edge cases in ML systems.
Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.
Original research-level math problems created by professional mathematicians. Problems are unpublished and cannot be memorized.
Graduate-level science questions written by PhD experts. Diamond subset contains questions where experts disagree, testing deep understanding.
Simple factual questions with verified correct answers. Tests accuracy of basic knowledge retrieval. Low scores indicate hallucination.
Tactical chess puzzles testing pattern recognition and multi-move calculation. Measures strategic reasoning ability.
- Typemultimodal
- Context262K tokens (~131 books)
- ReleasedApr 2026
- LicenseOpen Source
- StatusActive
- Cost / Message~$0.006
Frequently Asked Questions
Related Models
Qwen 3.5 Plus (hosted 397B-A17B) · Alibaba QwenGPT-3.5 Turbo (older v0613) · OpenAILlama 3.1 405B · MetaGemini 2.0 Flash (Dec 2024) · Google DeepMindQwen3.7 Plus · Alibaba QwenKey facts · as of 2026-10-05
- Kimi K2.6 by moonshotai. BenchGecko score 43.7, rank 181 of 312 scored models (normalized average of public benchmark scores).
- List price $0.95 input · $4.00 output per 1M tokens (as of 2026-10-05).
- Sold by 18 providers (as of 2026-10-05): Inceptron (int4) $0.47 in / $2.45 out · Chutes (int4) $0.50 in / $2.85 out · DigitalOcean $0.57 in / $2.40 out · Decart (fp4) $0.59 in / $2.47 out · StreamLake (fp8) $0.60 in / $2.52 out · and 13 more. Every provider
- Gecko Tests: Who Are You B (Knows who made it) · Tokenizer Tax E (92% more tokens outside English).
How to cite · data as of 2026-10-05
Kimi K2.6 · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/kimi-k2-6
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP