Moonshot AI, open weights
Kimi K2.6
Kimi K2.6 by Moonshot AI ranks 60th of 354 ranked models on the Noometry Index as of October 2026, with a score of 47.7. Its strongest category is writing & preference, where it ranks 26th. API pricing starts at $0.95 per million input tokens and $4 per million output tokens, with a 262K-token context window.
Last verified
Specifications
- Noometry rank
- #60 of 354
- Index score
- 47.7
- Evidence
- Confirmed 51 results
- Provider
- Moonshot AI
- Released
- April 20, 2026
- Weights
- Open weights
- Reasoning
- Yes
- Context window
- 262K
- Max output
- 262K
- Input price
- $0.95 / M
- Output price
- $4 / M
- Blended price
- $1.71 / M
- Output speed
- Not measured
- Value
- #138 of 219
- Knowledge cutoff
- January 2025
- Input
- text, image, video
- Hugging Face
- moonshotai/Kimi-K2.6
Category scores
Each category score combines every public result we have in that category.
- Coding 50.7
- Agentic & Tool Use 21.9
- Reasoning 40.5
- Math 57.0
- Knowledge 54.0
- Multimodal 31.6
- Multilingual 54.9
- Instruction Following 76.3
- Long Context 44.9
- Writing & Preference 68.5
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 50.7 | #43 | 5 |
| Agentic & Tool Use | 21.9 | #137 | 4 |
| Reasoning | 40.5 | #55 | 8 |
| Math | 57.0 | #41 | 6 |
| Knowledge | 54.0 | #54 | 4 |
| Multimodal | 31.6 | #103 | 3 |
| Multilingual | 54.9 | #37 | 1 |
| Instruction Following | 76.3 | #43 | 1 |
| Long Context | 44.9 | #52 | 1 |
| Writing & Preference | 68.5 | #26 | 5 |
Strengths and weaknesses
Categories where Kimi K2.6 places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Writing & Preference | 68.5 | +14.8 | #26 of 312, top 9% |
| Multilingual | 54.9 | +7.5 | #37 of 297, top 13% |
| Math | 57.0 | +20.5 | #41 of 327, top 13% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Agentic & Tool Use | 21.9 | −8.5 | #137 of 154, top 89% |
| Multimodal | 31.6 | −6.9 | #103 of 128, top 81% |
| Long Context | 44.9 | +4.0 | #52 of 296, top 18% |
Closest competitors
The models ranked just above and below Kimi K2.6. When scores are this close, price and speed are often the better way to choose.
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| SWE-bench Verified | 76.7% | #10 of 32, top 32% | Epoch AI | 2026-05-08 | |
| LMArena WebDev | 1509 | #45 of 113, top 40% | LMArena | 2026-10-08 | |
| SciCode | 53.5% | #32 of 121, top 27% | Epoch AI | ||
| WeirdML | 55.9% | #38 of 119, top 32% | Epoch AI | ||
| LMArena Coding | 1488 | #32 of 294, top 11% | LMArena | 2026-10-08 | |
| ALE-Bench | 1,093 | #38 of 105, top 37% | Epoch AI |
Agentic & Tool Use
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| OSWorld 2.0 | 4.6% | #7 of 9, top 78% | Epoch AI | ||
| ExploitBench | 18.4% | #6 of 9, top 67% | Epoch AI | ||
| GBAEval | 0.9% | #17 of 23, top 74% | Epoch AI | ||
| GDP.pdf | 12% | #32 of 36, top 89% | Epoch AI | ||
| Vending-Bench 2 | 6,205 | #18 of 60, top 30% | Epoch AI |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| NYT Connections (extended) | 87.2% | #26 of 91, top 29% | Lech Mazur benchmarks | ||
| CritPt | 8% | #54 of 134, top 41% | Epoch AI | ||
| Chess Puzzles | 26% | #40 of 129, top 32% | Epoch AI | 2026-05-07 | |
| EBR-Bench | 2.4% | #24 of 24, top 100% | Epoch AI | 2026-06-25 | |
| LMArena Hard Prompts | 1470 | #38 of 297, top 13% | LMArena | 2026-10-08 | |
| Mystery Game Puzzles | 18% | #47 of 74, top 64% | Epoch AI | 2026-07-17 | |
| Mystery Game Puzzles | 12% | none | Epoch AI | 2026-08-29 | |
| DTBench | 90.9% | #34 of 151, top 23% | Epoch AI | ||
| LMCA | 37.3% | #57 of 125, top 46% | Epoch AI | ||
| Epoch Capabilities Index | 151.05 | #49 of 213, top 24% | Epoch AI | 2026-04-20 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| FrontierMath (Tiers 1-3) | 57.2% | #40 of 81, top 50% | Epoch AI | 2026-06-10 | |
| FrontierMath Tier 4 | 25.6% | #37 of 63, top 59% | Epoch AI | 2026-06-10 | |
| MathArena Final-Answer Competitions | 72.9% | #11 of 29, top 38% | think | MathArena | |
| OTIS Mock AIME 2024-2025 | 96.1% | #30 of 173, top 18% | Epoch AI | 2026-05-02 | |
| ProofBench | 16% | #54 of 77, top 71% | Epoch AI | ||
| LMArena Math | 1475 | #32 of 285, top 12% | LMArena | 2026-10-08 | |
| FrontierMath (Feb 2025 set) | 39% | #11 of 68, top 17% | Epoch AI | 2026-05-07 | |
| FrontierMath Tier 4 (v1) | 14.6% | #16 of 55, top 30% | Epoch AI | 2026-05-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 90.8% | #32 of 186, top 18% | Epoch AI | 2026-05-01 | |
| SimpleQA Verified | 34.9% | #48 of 77, top 63% | Epoch AI | 2026-08-10 | |
| Vectara Hallucination Rate (lower is better) | 10.8% | #63 of 96, top 66% | Vectara Hallucination Leaderboard | ||
| LMArena Expert | 1491 | #30 of 273, top 11% | LMArena | 2026-10-08 |
Multimodal
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Vision | 1283 | #26 of 122, top 22% | LMArena | 2026-10-09 | |
| Blueprint-Bench 2 | 3.9% | #26 of 31, top 84% | Epoch AI | ||
| Furniture Assembly | 21.7% | #29 of 31, top 94% | Epoch AI | 2026-09-11 | |
| LMArena Document | 1451 | #22 of 38, top 58% | LMArena | 2026-09-13 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1446 | #37 of 297, top 13% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1521 | #25 of 285, top 9% | LMArena | 2026-10-08 | |
| LMArena French | 1471 | #37 of 223, top 17% | LMArena | 2026-10-08 | |
| LMArena German | 1450 | #44 of 231, top 20% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1443 | #28 of 211, top 14% | LMArena | 2026-10-08 | |
| LMArena Korean | 1427 | #30 of 213, top 15% | LMArena | 2026-10-08 | |
| LMArena Russian | 1446 | #49 of 283, top 18% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1464 | #32 of 226, top 15% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1451 | #41 of 298, top 14% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1468 | #34 of 291, top 12% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1455 | #38 of 297, top 13% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1434 | #47 of 295, top 16% | LMArena | 2026-10-08 | |
| EQ-Bench Creative Writing | 1725 | #28 of 115, top 25% | EQ-Bench | ||
| EQ-Bench 4 | 1202 | #17 of 28, top 61% | EQ-Bench | ||
| LMArena Multi-Turn | 1453 | #49 of 295, top 17% | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| azure | $0.95 | $4 | — | 2026-10-10 |
| deepinfra | $0.75 | $3.50 | $0.15 | 2026-10-10 |
| moonshot | $0.95 | $4 | $0.16 | 2026-10-10 |
| openrouter | $0.44 | $2.45 | $0.11 | 2026-10-10 |
Compare Kimi K2.6
- Kimi K2.6 vs Kimi K2.5
- Kimi K2.6 vs GLM-5.1
- Kimi K2.6 vs o3
- Kimi K2.6 vs Step 5 Preview
- Kimi K2.6 vs Qwen3.6 Plus
- Kimi K2.6 vs Inkling-Small
- Kimi K2.6 vs GPT-6 Astra
- Kimi K2.6 vs Claude Fable 5.1
- Kimi K2.6 vs Gemini 3.8 Flash
- Kimi K2.6 vs Grok 4.6
- Kimi K2.6 vs Qwen3.8 Max
- Kimi K2.6 vs GLM-5.3
- Kimi K2.6 vs Muse Spark 1.3
- Kimi K2.6 vs DeepSeek V4 Pro
Other Moonshot AI models
Frequently asked questions
How good is Kimi K2.6?
Kimi K2.6 by Moonshot AI ranks 60th of 354 ranked models on the Noometry Index as of October 2026, with a score of 47.7. Its strongest category is writing & preference, where it ranks 26th. API pricing starts at $0.95 per million input tokens and $4 per million output tokens, with a 262K-token context window.
How much does Kimi K2.6 cost?
Kimi K2.6 costs $0.95 per million input tokens and $4 per million output tokens on Moonshot AI's own API, with cached input at $0.16.
What is Kimi K2.6's context window?
Kimi K2.6 accepts up to 262K tokens of input and can write up to 262K tokens in one response.
Is Kimi K2.6 open source?
Yes. Kimi K2.6's weights are downloadable from Hugging Face (moonshotai/Kimi-K2.6); check the license for commercial terms.
What are Kimi K2.6's strengths and weaknesses?
Relative to other ranked models, Kimi K2.6 places best in writing & preference, multilingual, math and lowest in agentic & tool use, multimodal, long context.
What is Kimi K2.6 best at?
Its best category is writing & preference, where it ranks 26th on Noometry.