Model comparison

Kimi K2.7 Code vs Mistral Small

Kimi K2.7 Code is the stronger model overall, scoring 43.3 to 33.4 on the Noometry Index. Mistral Small costs 6.5× less per token, which makes it the better buy when Kimi K2.7 Code's lead doesn't matter for your workload.

Last verified . 5 shared benchmarks.

Kimi K2.7 Code Moonshot AI

43.3

Rank #94 Confirmed

Mistral Small Mistral AI

33.4

Rank #243 Confirmed

Summary

  • They share 5 benchmarks with published results for both. Kimi K2.7 Code scores higher in 4 categories and Mistral Small in 1 category; 5 gaps are clear of the uncertainty.
  • The widest gap is in math, where Kimi K2.7 Code leads 52.9 to 16.4.
  • The biggest single-benchmark swing is OTIS Mock AIME 2024-2025: 95.6% for Kimi K2.7 Code and 5.8% for Mistral Small.
  • Mistral Small is cheaper at $0.15 / $0.60 per million input/output tokens, against $0.95 / $4 for Kimi K2.7 Code.

Side by side

Kimi K2.7 Code and Mistral Small specifications
Kimi K2.7 CodeMistral Small
ProviderMoonshot AIMistral AI
Noometry Index43.333.4
Released2026-06-122024-02-26
WeightsOpenOpen
Context window262K262K
Max output262K256K
Input $ / M tokens$0.95$0.15
Output $ / M tokens$4$0.60
Results tracked1939

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Kimi K2.7 Code leads

Kimi K2.7 Code: 42.9 (#95), Mistral Small: 34.0 (#247)

Coding benchmarks
BenchmarkKimi K2.7 CodeMistral Small
SciCode47.5%26.5%
ALE-Bench886.23497.62
DeepSWE30.5%—
FrontierCode30.1%—
LMArena WebDev1473—
WeirdML54.1%—
BigCodeBench Instruct—36.1%
LiveBench Coding—36.2%
LMArena Coding—1362
BigCodeBench Complete—46.6%

Agentic & Tool Use Mistral Small leads

Kimi K2.7 Code: 24.0 (#122), Mistral Small: 28.1 (#93)

Agentic & Tool Use benchmarks
BenchmarkKimi K2.7 CodeMistral Small
APEX-Agents37.6%—
Berkeley Function Calling Leaderboard—37.1%
GBAEval0.9%—
Vending-Bench 25,083—

Reasoning Kimi K2.7 Code leads

Kimi K2.7 Code: 39.0 (#61), Mistral Small: 19.8 (#250)

Reasoning benchmarks
BenchmarkKimi K2.7 CodeMistral Small
CritPt10%0%
SimpleBench57.9%—
Kagi LLM Benchmark—37.8%
Chess Puzzles21%—
LiveBench Reasoning—44.8%
LMArena Hard Prompts—1335
DTBench—70.9%
LiveBench Data Analysis—53.7%
LMCA—20.6%
Surface Evolver Bench48.8%—
Epoch Capabilities Index149.97—
LiveBench—44%

Math Kimi K2.7 Code leads

Kimi K2.7 Code: 52.9 (#48), Mistral Small: 16.4 (#293)

Math benchmarks
BenchmarkKimi K2.7 CodeMistral Small
OTIS Mock AIME 2024-202595.6%5.8%
FrontierMath (Tiers 1-3)54%—
FrontierMath Tier 412.2%—
LiveBench Math—39.9%
LMArena Math—1341
MATH Level 5—46.8%

Knowledge Kimi K2.7 Code leads

Kimi K2.7 Code: 53.5 (#57), Mistral Small: 31.0 (#222)

Knowledge benchmarks
BenchmarkKimi K2.7 CodeMistral Small
GPQA Diamond87.9%47.5%
SimpleQA Verified36.5%—
Vectara Hallucination Rate—5.1%
LMArena Expert—1291
MMLU—68.7%

Multimodal Not comparable

Kimi K2.7 Code: —, Mistral Small: 33.5 (#96)

Multimodal benchmarks
BenchmarkKimi K2.7 CodeMistral Small
LMArena Vision—1142

Multilingual Not comparable

Kimi K2.7 Code: —, Mistral Small: 45.5 (#169)

Multilingual benchmarks
BenchmarkKimi K2.7 CodeMistral Small
LMArena Non-English—1315
LMArena Chinese—1340
LMArena French—1337
LMArena German—1340
LMArena Japanese—1275
LMArena Korean—1259
LMArena Russian—1324
LMArena Spanish—1346

Instruction Following Not comparable

Kimi K2.7 Code: —, Mistral Small: 66.4 (#209)

Instruction Following benchmarks
BenchmarkKimi K2.7 CodeMistral Small
LiveBench Instruction Following—63.7%
LMArena Instruction Following—1310

Long Context Not comparable

Kimi K2.7 Code: —, Mistral Small: 40.4 (#156)

Long Context benchmarks
BenchmarkKimi K2.7 CodeMistral Small
LMArena Longer Query—1327

Writing & Preference Not comparable

Kimi K2.7 Code: —, Mistral Small: 52.5 (#171)

Writing & Preference benchmarks
BenchmarkKimi K2.7 CodeMistral Small
LMArena Text—1338
LMArena Creative Writing—1305
LMArena Multi-Turn—1344
LiveBench Language—30.5%

Frequently asked questions

Is Kimi K2.7 Code better than Mistral Small?

Kimi K2.7 Code is the stronger model overall, scoring 43.3 to 33.4 on the Noometry Index. Mistral Small costs 6.5× less per token, which makes it the better buy when Kimi K2.7 Code's lead doesn't matter for your workload.

Which is cheaper, Kimi K2.7 Code or Mistral Small?

Mistral Small is cheaper. It lists at $0.15 per million input tokens and $0.60 per million output tokens; Kimi K2.7 Code lists at $0.95 and $4.

Is Kimi K2.7 Code or Mistral Small better for coding?

Kimi K2.7 Code scores higher on coding benchmarks: 42.9 versus 34.0 in the Noometry coding category.

Which has the bigger context window?

Both accept 262K tokens.

How many benchmarks do Kimi K2.7 Code and Mistral Small share?

5 benchmarks have published results for both models. Kimi K2.7 Code has 19 scored results on Noometry and Mistral Small has 39.

Related comparisons

Go deeper