Model comparison

Command A vs Command R+

Command A is the stronger model overall, scoring 36.5 to 32.4 on the Noometry Index.

Last verified . 20 shared benchmarks.

Command A Cohere

36.5

Rank #215 Confirmed

Command R+ Cohere

32.4

Rank #257 Confirmed

Summary

  • They share 20 benchmarks with published results for both. Command A scores higher in 7 categories and Command R+ in 1 category; 7 gaps are clear of the uncertainty.
  • The widest gap is in instruction following, where Command A leads 69.1 to 60.0.
  • The biggest single-benchmark swing is DTBench: 61.3% for Command A and 54.9% for Command R+.
  • Both cost about the same: $2.50 input and $10 output per million tokens.
  • Command A accepts more context: 256K tokens versus 128K.

Side by side

Command A and Command R+ specifications
Command ACommand R+
ProviderCohereCohere
Noometry Index36.532.4
Released2025-03-132024-08-30
WeightsOpenOpen
Context window256K128K
Max output8K4K
Input $ / M tokens$2.50$2.50
Output $ / M tokens$10$10
Results tracked2434

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Command R+ leads

Command A: 27.2 (#322), Command R+: 29.1 (#309)

Coding benchmarks
BenchmarkCommand ACommand R+
LMArena Coding13301187
Aider Polyglot12%—
BigCodeBench Instruct—33.8%
LiveBench Coding—19.1%
BigCodeBench Complete—41.9%
HumanEval+—56.7%
MBPP+—63.5%

Agentic & Tool Use Not comparable

Command A: 35.9 (#40), Command R+: —

Agentic & Tool Use benchmarks
BenchmarkCommand ACommand R+
Berkeley Function Calling Leaderboard57.1%—

Reasoning Command A leads

Command A: 18.3 (#283), Command R+: 9.2 (#344)

Reasoning benchmarks
BenchmarkCommand ACommand R+
LMArena Hard Prompts13261186
DTBench61.3%54.9%
LMCA10.3%5%
SimpleBench—17.4%
Kagi LLM Benchmark28.8%—
LiveBench Reasoning—24.8%
LiveBench Data Analysis—38.1%
Epoch Capabilities Index—119.34
LiveBench—31.8%

Math Command A leads

Command A: 36.2 (#171), Command R+: 28.9 (#242)

Math benchmarks
BenchmarkCommand ACommand R+
LMArena Math13001188
LiveBench Math—21.3%

Knowledge Too close to call

Command A: 37.1 (#159), Command R+: 36.4 (#169)

Knowledge benchmarks
BenchmarkCommand ACommand R+
Vectara Hallucination Rate9.3%6.9%
LMArena Expert12951174
MMLU—69.4%

Multilingual Command A leads

Command A: 45.3 (#170), Command R+: 38.6 (#227)

Multilingual benchmarks
BenchmarkCommand ACommand R+
LMArena Non-English13131216
LMArena Chinese13271226
LMArena French13511209
LMArena German13411216
LMArena Japanese12851166
LMArena Korean12851138
LMArena Russian13141227
LMArena Spanish13471189

Instruction Following Command A leads

Command A: 69.1 (#177), Command R+: 60.0 (#254)

Instruction Following benchmarks
BenchmarkCommand ACommand R+
LMArena Instruction Following13091197
LiveBench Instruction Following—57.6%

Long Context Command A leads

Command A: 40.6 (#151), Command R+: 37.3 (#219)

Long Context benchmarks
BenchmarkCommand ACommand R+
LMArena Longer Query13341230

Writing & Preference Command A leads

Command A: 47.6 (#208), Command R+: 43.5 (#228)

Writing & Preference benchmarks
BenchmarkCommand ACommand R+
LMArena Text13311229
LMArena Creative Writing13191235
LMArena Multi-Turn13391213
EQ-Bench Creative Writing1145—
LiveBench Language—29.7%

Frequently asked questions

Is Command A better than Command R+?

Command A is the stronger model overall, scoring 36.5 to 32.4 on the Noometry Index.

Which is cheaper, Command A or Command R+?

Command R+ is cheaper. It lists at $2.50 per million input tokens and $10 per million output tokens; Command A lists at $2.50 and $10.

Is Command A or Command R+ better for coding?

Command R+ scores higher on coding benchmarks: 29.1 versus 27.2 in the Noometry coding category.

Which has the bigger context window?

Command A does, with 256K tokens against 128K.

How many benchmarks do Command A and Command R+ share?

20 benchmarks have published results for both models. Command A has 24 scored results on Noometry and Command R+ has 34.

Related comparisons

Go deeper