Model comparison

Command R vs Gemini 1.5 Pro (May 2024)

Command R and Gemini 1.5 Pro (May 2024) score almost the same on the Noometry Index (31.4 vs 32.1), so choose on price, context window or the category you care about most.

Last verified . 21 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Gemini 1.5 Pro (May 2024) Google

32.1

Rank #261 Confirmed

Summary

  • They share 21 benchmarks with published results for both. Command R scores higher in 3 categories and Gemini 1.5 Pro (May 2024) in 5 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Gemini 1.5 Pro (May 2024) leads 52.4 to 38.2.
  • The biggest single-benchmark swing is DTBench: 46.4% for Command R and 59% for Gemini 1.5 Pro (May 2024).
  • Command R has downloadable open weights; the other is API-only.

Side by side

Command R and Gemini 1.5 Pro (May 2024) specifications
Command RGemini 1.5 Pro (May 2024)
ProviderCohereGoogle
Noometry Index31.432.1
Released2024-08-302024-02-15
WeightsOpenProprietary
Context window128K—
Max output4K—
Input $ / M tokens$0.15—
Output $ / M tokens$0.60—
Results tracked2945

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 1.5 Pro (May 2024) leads

Command R: 29.3 (#306), Gemini 1.5 Pro (May 2024): 34.2 (#241)

Coding benchmarks
BenchmarkCommand RGemini 1.5 Pro (May 2024)
BigCodeBench Instruct37.1%43.8%
LMArena Coding11691294
BigCodeBench Complete45.2%57.5%
WeirdML—22.2%
LiveBench Coding17.9%—
CadEval—34%
HumanEval+—79.3%
MBPP+—74.6%

Agentic & Tool Use Not comparable

Command R: —, Gemini 1.5 Pro (May 2024): 17.9 (#145)

Agentic & Tool Use benchmarks
BenchmarkCommand RGemini 1.5 Pro (May 2024)
TheAgentCompany—3.4%
Cybench—7.5%
BALROG—21%

Reasoning Command R leads

Command R: 13.8 (#331), Gemini 1.5 Pro (May 2024): 12.3 (#338)

Reasoning benchmarks
BenchmarkCommand RGemini 1.5 Pro (May 2024)
LMArena Hard Prompts11641296
DTBench46.4%59%
ARC-AGI-2—0.8%
SimpleBench—27.1%
LiveBench Reasoning21.9%—
LiveBench Data Analysis33.3%—
LMCA9.2%—
BIG-Bench Hard—89.2%
Epoch Capabilities Index—131.73
ForecastBench—58.4
LiveBench27.5%—

Math Command R leads

Command R: 28.0 (#246), Gemini 1.5 Pro (May 2024): 25.8 (#266)

Math benchmarks
BenchmarkCommand RGemini 1.5 Pro (May 2024)
LMArena Math11551315
OTIS Mock AIME 2024-2025—23.1%
Omni-MATH—36.4%
LiveBench Math19.4%—
MATH Level 5—70.4%

Knowledge Command R leads

Command R: 31.0 (#221), Gemini 1.5 Pro (May 2024): 29.4 (#239)

Knowledge benchmarks
BenchmarkCommand RGemini 1.5 Pro (May 2024)
LMArena Expert11381279
MMLU65.2%86.9%
GPQA Diamond—57.2%
Humanity's Last Exam—4.6%
MMLU-Pro—73.7%
Confabulations—13.5%
GPQA (HELM)—53.4%

Multimodal Not comparable

Command R: —, Gemini 1.5 Pro (May 2024): 36.8 (#77)

Multimodal benchmarks
BenchmarkCommand RGemini 1.5 Pro (May 2024)
LMArena Vision—1161
Video-MME—75%

Multilingual Gemini 1.5 Pro (May 2024) leads

Command R: 35.7 (#245), Gemini 1.5 Pro (May 2024): 45.3 (#174)

Multilingual benchmarks
BenchmarkCommand RGemini 1.5 Pro (May 2024)
LMArena Non-English11741312
LMArena Chinese11821331
LMArena French11621302
LMArena German11761286
LMArena Japanese11431292
LMArena Korean11631298
LMArena Russian11741320
LMArena Spanish11511311

Instruction Following Gemini 1.5 Pro (May 2024) leads

Command R: 58.1 (#261), Gemini 1.5 Pro (May 2024): 68.6 (#185)

Instruction Following benchmarks
BenchmarkCommand RGemini 1.5 Pro (May 2024)
LMArena Instruction Following11671297
LiveBench Instruction Following55.6%—
IFEval—83.7%

Long Context Gemini 1.5 Pro (May 2024) leads

Command R: 36.3 (#231), Gemini 1.5 Pro (May 2024): 39.8 (#169)

Long Context benchmarks
BenchmarkCommand RGemini 1.5 Pro (May 2024)
LMArena Longer Query11981308

Writing & Preference Gemini 1.5 Pro (May 2024) leads

Command R: 38.2 (#254), Gemini 1.5 Pro (May 2024): 52.4 (#172)

Writing & Preference benchmarks
BenchmarkCommand RGemini 1.5 Pro (May 2024)
LMArena Text11871319
LMArena Creative Writing11701333
LMArena Multi-Turn11631296
WildBench—81.3%
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Gemini 1.5 Pro (May 2024)?

Command R and Gemini 1.5 Pro (May 2024) score almost the same on the Noometry Index (31.4 vs 32.1), so choose on price, context window or the category you care about most.

Is Command R or Gemini 1.5 Pro (May 2024) better for coding?

Gemini 1.5 Pro (May 2024) scores higher on coding benchmarks: 34.2 versus 29.3 in the Noometry coding category.

How many benchmarks do Command R and Gemini 1.5 Pro (May 2024) share?

21 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Gemini 1.5 Pro (May 2024) has 45.

Related comparisons

Go deeper