Model comparison

Command R vs Kimi K2.5 Instant

Kimi K2.5 Instant is the stronger model overall, scoring 43.6 to 31.4 on the Noometry Index.

Last verified . 16 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Kimi K2.5 Instant Moonshot AI

43.6

Rank #89 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Command R scores higher in 0 categories and Kimi K2.5 Instant in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Kimi K2.5 Instant leads 60.6 to 38.2.

Side by side

Command R and Kimi K2.5 Instant specifications
Command RKimi K2.5 Instant
ProviderCohereMoonshot AI
Noometry Index31.443.6
Released2024-08-30—
WeightsOpenOpen
Context window128K—
Max output4K—
Input $ / M tokens$0.15—
Output $ / M tokens$0.60—
Results tracked2918

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Kimi K2.5 Instant leads

Command R: 29.3 (#306), Kimi K2.5 Instant: 42.6 (#97)

Coding benchmarks
BenchmarkCommand RKimi K2.5 Instant
LMArena Coding11691484
LMArena WebDev—1404
BigCodeBench Instruct37.1%—
LiveBench Coding17.9%—
BigCodeBench Complete45.2%—

Reasoning Kimi K2.5 Instant leads

Command R: 13.8 (#331), Kimi K2.5 Instant: 29.7 (#90)

Reasoning benchmarks
BenchmarkCommand RKimi K2.5 Instant
LMArena Hard Prompts11641443
LiveBench Reasoning21.9%—
DTBench46.4%—
LiveBench Data Analysis33.3%—
LMCA9.2%—
LiveBench27.5%—

Math Kimi K2.5 Instant leads

Command R: 28.0 (#246), Kimi K2.5 Instant: 39.4 (#105)

Math benchmarks
BenchmarkCommand RKimi K2.5 Instant
LMArena Math11551442
LiveBench Math19.4%—

Knowledge Kimi K2.5 Instant leads

Command R: 31.0 (#221), Kimi K2.5 Instant: 40.2 (#123)

Knowledge benchmarks
BenchmarkCommand RKimi K2.5 Instant
LMArena Expert11381440
MMLU65.2%—

Multimodal Not comparable

Command R: —, Kimi K2.5 Instant: 40.2 (#50)

Multimodal benchmarks
BenchmarkCommand RKimi K2.5 Instant
LMArena Vision—1254

Multilingual Kimi K2.5 Instant leads

Command R: 35.7 (#245), Kimi K2.5 Instant: 52.0 (#94)

Multilingual benchmarks
BenchmarkCommand RKimi K2.5 Instant
LMArena Non-English11741406
LMArena Chinese11821449
LMArena French11621403
LMArena German11761413
LMArena Korean11631378
LMArena Russian11741404
LMArena Spanish11511447
LMArena Japanese1143—

Instruction Following Kimi K2.5 Instant leads

Command R: 58.1 (#261), Kimi K2.5 Instant: 75.3 (#65)

Instruction Following benchmarks
BenchmarkCommand RKimi K2.5 Instant
LMArena Instruction Following11671430
LiveBench Instruction Following55.6%—

Long Context Kimi K2.5 Instant leads

Command R: 36.3 (#231), Kimi K2.5 Instant: 43.9 (#83)

Long Context benchmarks
BenchmarkCommand RKimi K2.5 Instant
LMArena Longer Query11981435

Writing & Preference Kimi K2.5 Instant leads

Command R: 38.2 (#254), Kimi K2.5 Instant: 60.6 (#95)

Writing & Preference benchmarks
BenchmarkCommand RKimi K2.5 Instant
LMArena Text11871420
LMArena Creative Writing11701381
LMArena Multi-Turn11631427
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Kimi K2.5 Instant?

Kimi K2.5 Instant is the stronger model overall, scoring 43.6 to 31.4 on the Noometry Index.

Is Command R or Kimi K2.5 Instant better for coding?

Kimi K2.5 Instant scores higher on coding benchmarks: 42.6 versus 29.3 in the Noometry coding category.

How many benchmarks do Command R and Kimi K2.5 Instant share?

16 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Kimi K2.5 Instant has 18.

Related comparisons

Go deeper