Model comparison

Command R vs Gemma 1.1 2b IT

Command R is the stronger model overall, scoring 31.4 to 29.3 on the Noometry Index.

Last verified . 14 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Gemma 1.1 2b IT Google

29.3

Rank #313 Confirmed

Summary

  • They share 14 benchmarks with published results for both. Command R scores higher in 5 categories and Gemma 1.1 2b IT in 3 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Command R leads 38.2 to 25.1.

Side by side

Command R and Gemma 1.1 2b IT specifications
Command RGemma 1.1 2b IT
ProviderCohereGoogle
Noometry Index31.429.3
Released2024-08-30—
WeightsOpenOpen
Context window128K—
Max output4K—
Input $ / M tokens$0.15—
Output $ / M tokens$0.60—
Results tracked2916

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Command R: 29.3 (#306), Gemma 1.1 2b IT: 30.1 (#299)

Coding benchmarks
BenchmarkCommand RGemma 1.1 2b IT
LMArena Coding11691034
BigCodeBench Instruct37.1%—
LiveBench Coding17.9%—
BigCodeBench Complete45.2%—
HumanEval+—17.7%
MBPP+—23.3%

Reasoning Gemma 1.1 2b IT leads

Command R: 13.8 (#331), Gemma 1.1 2b IT: 19.1 (#270)

Reasoning benchmarks
BenchmarkCommand RGemma 1.1 2b IT
LMArena Hard Prompts11641005
LiveBench Reasoning21.9%—
DTBench46.4%—
LiveBench Data Analysis33.3%—
LMCA9.2%—
LiveBench27.5%—

Math Gemma 1.1 2b IT leads

Command R: 28.0 (#246), Gemma 1.1 2b IT: 30.8 (#232)

Math benchmarks
BenchmarkCommand RGemma 1.1 2b IT
LMArena Math11551047
LiveBench Math19.4%—

Knowledge Command R leads

Command R: 31.0 (#221), Gemma 1.1 2b IT: 26.5 (#258)

Knowledge benchmarks
BenchmarkCommand RGemma 1.1 2b IT
LMArena Expert1138970
MMLU65.2%—

Multilingual Command R leads

Command R: 35.7 (#245), Gemma 1.1 2b IT: 24.6 (#289)

Multilingual benchmarks
BenchmarkCommand RGemma 1.1 2b IT
LMArena Non-English1174988
LMArena Chinese11821012
LMArena German1176944
LMArena Korean1163899
LMArena Russian1174990
LMArena French1162—
LMArena Japanese1143—
LMArena Spanish1151—

Instruction Following Command R leads

Command R: 58.1 (#261), Gemma 1.1 2b IT: 49.9 (#299)

Instruction Following benchmarks
BenchmarkCommand RGemma 1.1 2b IT
LMArena Instruction Following1167992
LiveBench Instruction Following55.6%—

Long Context Command R leads

Command R: 36.3 (#231), Gemma 1.1 2b IT: 30.6 (#286)

Long Context benchmarks
BenchmarkCommand RGemma 1.1 2b IT
LMArena Longer Query11981003

Writing & Preference Command R leads

Command R: 38.2 (#254), Gemma 1.1 2b IT: 25.1 (#306)

Writing & Preference benchmarks
BenchmarkCommand RGemma 1.1 2b IT
LMArena Text11871022
LMArena Creative Writing1170998
LMArena Multi-Turn1163959
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Gemma 1.1 2b IT?

Command R is the stronger model overall, scoring 31.4 to 29.3 on the Noometry Index.

Is Command R or Gemma 1.1 2b IT better for coding?

They score almost the same on coding (29.3 vs 30.1); test both on your own repository before choosing.

How many benchmarks do Command R and Gemma 1.1 2b IT share?

14 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Gemma 1.1 2b IT has 16.

Related comparisons

Go deeper