Model comparison

Command R vs Gemma 3n E4b IT

Gemma 3n E4b IT is the stronger model overall, scoring 37.3 to 31.4 on the Noometry Index.

Last verified . 17 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Gemma 3n E4b IT Google

37.3

Rank #206 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Command R scores higher in 0 categories and Gemma 3n E4b IT in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Gemma 3n E4b IT leads 50.1 to 38.2.

Side by side

Command R and Gemma 3n E4b IT specifications
Command RGemma 3n E4b IT
ProviderCohereGoogle
Noometry Index31.437.3
Released2024-08-30—
WeightsOpenOpen
Context window128K—
Max output4K—
Input $ / M tokens$0.15—
Output $ / M tokens$0.60—
Results tracked2918

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemma 3n E4b IT leads

Command R: 29.3 (#306), Gemma 3n E4b IT: 37.0 (#198)

Coding benchmarks
BenchmarkCommand RGemma 3n E4b IT
LMArena Coding11691268
BigCodeBench Instruct37.1%—
LiveBench Coding17.9%—
BigCodeBench Complete45.2%—

Reasoning Gemma 3n E4b IT leads

Command R: 13.8 (#331), Gemma 3n E4b IT: 19.9 (#247)

Reasoning benchmarks
BenchmarkCommand RGemma 3n E4b IT
LMArena Hard Prompts11641284
Kagi LLM Benchmark—31.5%
LiveBench Reasoning21.9%—
DTBench46.4%—
LiveBench Data Analysis33.3%—
LMCA9.2%—
LiveBench27.5%—

Math Gemma 3n E4b IT leads

Command R: 28.0 (#246), Gemma 3n E4b IT: 35.1 (#188)

Math benchmarks
BenchmarkCommand RGemma 3n E4b IT
LMArena Math11551251
LiveBench Math19.4%—

Knowledge Gemma 3n E4b IT leads

Command R: 31.0 (#221), Gemma 3n E4b IT: 34.2 (#198)

Knowledge benchmarks
BenchmarkCommand RGemma 3n E4b IT
LMArena Expert11381246
MMLU65.2%—

Multilingual Gemma 3n E4b IT leads

Command R: 35.7 (#245), Gemma 3n E4b IT: 43.4 (#183)

Multilingual benchmarks
BenchmarkCommand RGemma 3n E4b IT
LMArena Non-English11741285
LMArena Chinese11821309
LMArena French11621330
LMArena German11761311
LMArena Japanese11431272
LMArena Korean11631259
LMArena Russian11741288
LMArena Spanish11511305

Instruction Following Gemma 3n E4b IT leads

Command R: 58.1 (#261), Gemma 3n E4b IT: 66.1 (#210)

Instruction Following benchmarks
BenchmarkCommand RGemma 3n E4b IT
LMArena Instruction Following11671255
LiveBench Instruction Following55.6%—

Long Context Gemma 3n E4b IT leads

Command R: 36.3 (#231), Gemma 3n E4b IT: 38.7 (#191)

Long Context benchmarks
BenchmarkCommand RGemma 3n E4b IT
LMArena Longer Query11981276

Writing & Preference Gemma 3n E4b IT leads

Command R: 38.2 (#254), Gemma 3n E4b IT: 50.1 (#186)

Writing & Preference benchmarks
BenchmarkCommand RGemma 3n E4b IT
LMArena Text11871306
LMArena Creative Writing11701287
LMArena Multi-Turn11631276
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Gemma 3n E4b IT?

Gemma 3n E4b IT is the stronger model overall, scoring 37.3 to 31.4 on the Noometry Index.

Is Command R or Gemma 3n E4b IT better for coding?

Gemma 3n E4b IT scores higher on coding benchmarks: 37.0 versus 29.3 in the Noometry coding category.

How many benchmarks do Command R and Gemma 3n E4b IT share?

17 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Gemma 3n E4b IT has 18.

Related comparisons

Go deeper