Model comparison

Kimi K2.7 Code vs Mistral Medium 3.5

Kimi K2.7 Code is the stronger model overall, scoring 43.3 to 40.2 on the Noometry Index.

Last verified . 2 shared benchmarks.

Kimi K2.7 Code Moonshot AI

43.3

Rank #94 Confirmed

Mistral Medium 3.5 Mistral AI

40.2

Rank #152 Confirmed

Summary

  • They share 2 benchmarks with published results for both. Kimi K2.7 Code scores higher in 4 categories and Mistral Medium 3.5 in 0 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Kimi K2.7 Code leads 39.0 to 17.3.
  • Kimi K2.7 Code is cheaper at $0.95 / $4 per million input/output tokens, against $1.50 / $7.50 for Mistral Medium 3.5.

Side by side

Kimi K2.7 Code and Mistral Medium 3.5 specifications
Kimi K2.7 CodeMistral Medium 3.5
ProviderMoonshot AIMistral AI
Noometry Index43.340.2
Released2026-06-12—
WeightsOpenOpen
Context window262K262K
Max output262K210K
Input $ / M tokens$0.95$1.50
Output $ / M tokens$4$7.50
Results tracked1922

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Kimi K2.7 Code leads

Kimi K2.7 Code: 42.9 (#95), Mistral Medium 3.5: 36.0 (#213)

Coding benchmarks
BenchmarkKimi K2.7 CodeMistral Medium 3.5
LMArena WebDev14731264
DeepSWE30.5%—
FrontierCode30.1%—
SciCode47.5%—
WeirdML54.1%—
LMArena Coding—1461
ALE-Bench886.23—

Agentic & Tool Use Not comparable

Kimi K2.7 Code: 24.0 (#122), Mistral Medium 3.5: —

Agentic & Tool Use benchmarks
BenchmarkKimi K2.7 CodeMistral Medium 3.5
APEX-Agents37.6%—
GBAEval0.9%—
Vending-Bench 25,083—

Reasoning Kimi K2.7 Code leads

Kimi K2.7 Code: 39.0 (#61), Mistral Medium 3.5: 17.3 (#295)

Reasoning benchmarks
BenchmarkKimi K2.7 CodeMistral Medium 3.5
Epoch Capabilities Index149.97141.35
SimpleBench57.9%—
Kagi LLM Benchmark—41.4%
NYT Connections (extended)—12.9%
CritPt10%—
Chess Puzzles21%—
LMArena Hard Prompts—1436
Surface Evolver Bench48.8%—

Math Kimi K2.7 Code leads

Kimi K2.7 Code: 52.9 (#48), Mistral Medium 3.5: 39.1 (#113)

Math benchmarks
BenchmarkKimi K2.7 CodeMistral Medium 3.5
FrontierMath (Tiers 1-3)54%—
FrontierMath Tier 412.2%—
OTIS Mock AIME 2024-202595.6%—
LMArena Math—1431

Knowledge Kimi K2.7 Code leads

Kimi K2.7 Code: 53.5 (#57), Mistral Medium 3.5: 40.0 (#126)

Knowledge benchmarks
BenchmarkKimi K2.7 CodeMistral Medium 3.5
GPQA Diamond87.9%—
SimpleQA Verified36.5%—
LMArena Expert—1432

Multimodal Not comparable

Kimi K2.7 Code: —, Mistral Medium 3.5: 38.3 (#65)

Multimodal benchmarks
BenchmarkKimi K2.7 CodeMistral Medium 3.5
LMArena Vision—1223

Multilingual Not comparable

Kimi K2.7 Code: —, Mistral Medium 3.5: 51.9 (#100)

Multilingual benchmarks
BenchmarkKimi K2.7 CodeMistral Medium 3.5
LMArena Non-English—1404
LMArena Chinese—1442
LMArena French—1448
LMArena German—1451
LMArena Korean—1385
LMArena Russian—1395
LMArena Spanish—1409

Instruction Following Not comparable

Kimi K2.7 Code: —, Mistral Medium 3.5: 74.6 (#90)

Instruction Following benchmarks
BenchmarkKimi K2.7 CodeMistral Medium 3.5
LMArena Instruction Following—1415

Long Context Not comparable

Kimi K2.7 Code: —, Mistral Medium 3.5: 43.2 (#103)

Long Context benchmarks
BenchmarkKimi K2.7 CodeMistral Medium 3.5
LMArena Longer Query—1415

Writing & Preference Not comparable

Kimi K2.7 Code: —, Mistral Medium 3.5: 58.5 (#117)

Writing & Preference benchmarks
BenchmarkKimi K2.7 CodeMistral Medium 3.5
LMArena Text—1421
LMArena Creative Writing—1374
EQ-Bench 4—993
LMArena Multi-Turn—1423

Frequently asked questions

Is Kimi K2.7 Code better than Mistral Medium 3.5?

Kimi K2.7 Code is the stronger model overall, scoring 43.3 to 40.2 on the Noometry Index.

Which is cheaper, Kimi K2.7 Code or Mistral Medium 3.5?

Kimi K2.7 Code is cheaper. It lists at $0.95 per million input tokens and $4 per million output tokens; Mistral Medium 3.5 lists at $1.50 and $7.50.

Is Kimi K2.7 Code or Mistral Medium 3.5 better for coding?

Kimi K2.7 Code scores higher on coding benchmarks: 42.9 versus 36.0 in the Noometry coding category.

Which has the bigger context window?

Both accept 262K tokens.

How many benchmarks do Kimi K2.7 Code and Mistral Medium 3.5 share?

2 benchmarks have published results for both models. Kimi K2.7 Code has 19 scored results on Noometry and Mistral Medium 3.5 has 22.

Related comparisons

Go deeper