Model comparison

Kimi K2.7 Code vs Wizardlm 70b

Kimi K2.7 Code is the stronger model overall, scoring 43.3 to 33.0 on the Noometry Index.

Last verified . 0 shared benchmarks.

Kimi K2.7 Code Moonshot AI

43.3

Rank #94 Confirmed

Wizardlm 70b Microsoft

33.0

Rank #249 Confirmed

Summary

  • The widest gap is in math, where Kimi K2.7 Code leads 52.9 to 32.2.

Side by side

Kimi K2.7 Code and Wizardlm 70b specifications
Kimi K2.7 CodeWizardlm 70b
ProviderMoonshot AIMicrosoft
Noometry Index43.333.0
Released2026-06-12—
WeightsOpenOpen
Context window262K—
Max output262K—
Input $ / M tokens$0.95—
Output $ / M tokens$4—
Results tracked1912

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Kimi K2.7 Code leads

Kimi K2.7 Code: 42.9 (#95), Wizardlm 70b: 31.4 (#285)

Coding benchmarks
BenchmarkKimi K2.7 CodeWizardlm 70b
DeepSWE30.5%—
FrontierCode30.1%—
LMArena WebDev1473—
SciCode47.5%—
WeirdML54.1%—
LMArena Coding—1081
ALE-Bench886.23—

Agentic & Tool Use Not comparable

Kimi K2.7 Code: 24.0 (#122), Wizardlm 70b: —

Agentic & Tool Use benchmarks
BenchmarkKimi K2.7 CodeWizardlm 70b
APEX-Agents37.6%—
GBAEval0.9%—
Vending-Bench 25,083—

Reasoning Kimi K2.7 Code leads

Kimi K2.7 Code: 39.0 (#61), Wizardlm 70b: 20.7 (#234)

Reasoning benchmarks
BenchmarkKimi K2.7 CodeWizardlm 70b
SimpleBench57.9%—
CritPt10%—
Chess Puzzles21%—
LMArena Hard Prompts—1079
Surface Evolver Bench48.8%—
Epoch Capabilities Index149.97—

Math Kimi K2.7 Code leads

Kimi K2.7 Code: 52.9 (#48), Wizardlm 70b: 32.2 (#218)

Math benchmarks
BenchmarkKimi K2.7 CodeWizardlm 70b
FrontierMath (Tiers 1-3)54%—
FrontierMath Tier 412.2%—
OTIS Mock AIME 2024-202595.6%—
LMArena Math—1116

Knowledge Not comparable

Kimi K2.7 Code: 53.5 (#57), Wizardlm 70b: —

Knowledge benchmarks
BenchmarkKimi K2.7 CodeWizardlm 70b
GPQA Diamond87.9%—
SimpleQA Verified36.5%—

Multilingual Not comparable

Kimi K2.7 Code: —, Wizardlm 70b: 29.6 (#265)

Multilingual benchmarks
BenchmarkKimi K2.7 CodeWizardlm 70b
LMArena Non-English—1078
LMArena Chinese—1052
LMArena German—1083
LMArena Russian—1155

Instruction Following Not comparable

Kimi K2.7 Code: —, Wizardlm 70b: 56.3 (#273)

Instruction Following benchmarks
BenchmarkKimi K2.7 CodeWizardlm 70b
LMArena Instruction Following—1093

Long Context Not comparable

Kimi K2.7 Code: —, Wizardlm 70b: 33.3 (#263)

Long Context benchmarks
BenchmarkKimi K2.7 CodeWizardlm 70b
LMArena Longer Query—1097

Writing & Preference Not comparable

Kimi K2.7 Code: —, Wizardlm 70b: 34.8 (#269)

Writing & Preference benchmarks
BenchmarkKimi K2.7 CodeWizardlm 70b
LMArena Text—1120
LMArena Creative Writing—1149
LMArena Multi-Turn—1108

Frequently asked questions

Is Kimi K2.7 Code better than Wizardlm 70b?

Kimi K2.7 Code is the stronger model overall, scoring 43.3 to 33.0 on the Noometry Index.

Is Kimi K2.7 Code or Wizardlm 70b better for coding?

Kimi K2.7 Code scores higher on coding benchmarks: 42.9 versus 31.4 in the Noometry coding category.

How many benchmarks do Kimi K2.7 Code and Wizardlm 70b share?

0 benchmarks have published results for both models. Kimi K2.7 Code has 19 scored results on Noometry and Wizardlm 70b has 12.

Related comparisons

Go deeper