Moonshot AI, open weights
Kimi K2 Thinking Turbo
Kimi K2 Thinking Turbo by Moonshot AI ranks 70th of 354 ranked models on the Noometry Index as of October 2026, with a score of 45.8. Its strongest category is math, where it ranks 68th.
Last verified
Specifications
- Noometry rank
- #70 of 354
- Index score
- 45.8
- Evidence
- Confirmed 21 results
- Provider
- Moonshot AI
- Released
- November 6, 2025
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 38.4
- Reasoning 32.5
- Math 47.4
- Knowledge 50.9
- Multilingual 51.4
- Instruction Following 74.0
- Long Context 43.2
- Writing & Preference 60.0
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 38.4 | #178 | 2 |
| Reasoning | 32.5 | #75 | 2 |
| Math | 47.4 | #68 | 2 |
| Knowledge | 50.9 | #69 | 2 |
| Multilingual | 51.4 | #109 | 1 |
| Instruction Following | 74.0 | #109 | 1 |
| Long Context | 43.2 | #102 | 1 |
| Writing & Preference | 60.0 | #104 | 3 |
Strengths and weaknesses
Categories where Kimi K2 Thinking Turbo places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Coding | 38.4 | −0.4 | #178 of 340, top 53% |
| Multilingual | 51.4 | +4.0 | #109 of 297, top 37% |
| Instruction Following | 74.0 | +2.7 | #109 of 305, top 36% |
Closest competitors
The models ranked just above and below Kimi K2 Thinking Turbo. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| GLM-5 | #66 | 46.1 | $1.55 | 23 | Compare |
| Qwen3.5 397B-A17B | #67 | 46.0 | $1.35 | 9 | Compare |
| Qwen3.8 27B | #68 | 46.0 | $1.11 | — | Compare |
| GPT-5.3 Codex | #69 | 45.8 | $4.81 | — | Compare |
| Qwen3.5 Max Preview | #71 | 45.3 | — | — | Compare |
| Qwen3.7 Plus | #72 | 45.3 | $0.70 | — | Compare |
| Hy4 preview | #73 | 45.3 | $1.13 | — | Compare |
| MiMo-V2.5-Pro | #74 | 45.2 | $0.54 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena WebDev | 1322 | #95 of 113, top 85% | LMArena | 2026-10-08 | |
| LMArena Coding | 1454 | #82 of 294, top 28% | LMArena | 2026-10-08 |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Chess Puzzles | 20% | #59 of 129, top 46% | Epoch AI | 2025-12-10 | |
| LMArena Hard Prompts | 1428 | #92 of 297, top 31% | LMArena | 2026-10-08 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| OTIS Mock AIME 2024-2025 | 83.1% | #76 of 173, top 44% | Epoch AI | 2025-11-13 | |
| LMArena Math | 1429 | #89 of 285, top 32% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 84.2% | #67 of 186, top 37% | Epoch AI | 2025-11-11 | |
| LMArena Expert | 1439 | #88 of 273, top 33% | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1398 | #109 of 297, top 37% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1456 | #98 of 285, top 35% | LMArena | 2026-10-08 | |
| LMArena French | 1425 | #96 of 223, top 44% | LMArena | 2026-10-08 | |
| LMArena German | 1390 | #104 of 231, top 46% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1357 | #93 of 211, top 45% | LMArena | 2026-10-08 | |
| LMArena Korean | 1330 | #114 of 213, top 54% | LMArena | 2026-10-08 | |
| LMArena Russian | 1391 | #117 of 283, top 42% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1406 | #104 of 226, top 47% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1403 | #103 of 298, top 35% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1415 | #97 of 291, top 34% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1415 | #105 of 297, top 36% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1374 | #108 of 295, top 37% | LMArena | 2026-10-08 | |
| LMArena Multi-Turn | 1414 | #103 of 295, top 35% | LMArena | 2026-10-08 |
Compare Kimi K2 Thinking Turbo
- Kimi K2 Thinking Turbo vs GPT-5.3 Codex
- Kimi K2 Thinking Turbo vs Qwen3.5 Max Preview
- Kimi K2 Thinking Turbo vs Qwen3.8 27B
- Kimi K2 Thinking Turbo vs Qwen3.7 Plus
- Kimi K2 Thinking Turbo vs Qwen3.5 397B-A17B
- Kimi K2 Thinking Turbo vs Hy4 preview
- Kimi K2 Thinking Turbo vs GPT-6 Astra
- Kimi K2 Thinking Turbo vs Claude Fable 5.1
- Kimi K2 Thinking Turbo vs Gemini 3.8 Flash
- Kimi K2 Thinking Turbo vs Grok 4.6
- Kimi K2 Thinking Turbo vs Qwen3.8 Max
- Kimi K2 Thinking Turbo vs GLM-5.3
- Kimi K2 Thinking Turbo vs Muse Spark 1.3
- Kimi K2 Thinking Turbo vs DeepSeek V4 Pro
Other Moonshot AI models
Frequently asked questions
How good is Kimi K2 Thinking Turbo?
Kimi K2 Thinking Turbo by Moonshot AI ranks 70th of 354 ranked models on the Noometry Index as of October 2026, with a score of 45.8. Its strongest category is math, where it ranks 68th.
Is Kimi K2 Thinking Turbo open source?
Yes. Kimi K2 Thinking Turbo's weights are downloadable; check the license for commercial terms.
What are Kimi K2 Thinking Turbo's strengths and weaknesses?
Relative to other ranked models, Kimi K2 Thinking Turbo places best in math, reasoning, knowledge and lowest in coding, multilingual, instruction following.
What is Kimi K2 Thinking Turbo best at?
Its best category is math, where it ranks 68th on Noometry.