Model comparison

Command A vs Command R

Command A is the stronger model overall, scoring 36.5 to 31.4 on the Noometry Index. Command R costs 17× less per token, which makes it the better buy when Command A's lead doesn't matter for your workload.

Last verified . 19 shared benchmarks.

Command A Cohere

36.5

Rank #215 Confirmed

Command R Cohere

31.4

Rank #272 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Command A scores higher in 7 categories and Command R in 1 category; 8 gaps are clear of the uncertainty.
  • The widest gap is in instruction following, where Command A leads 69.1 to 58.1.
  • The biggest single-benchmark swing is DTBench: 61.3% for Command A and 46.4% for Command R.
  • Command R is cheaper at $0.15 / $0.60 per million input/output tokens, against $2.50 / $10 for Command A.
  • Command A accepts more context: 256K tokens versus 128K.

Side by side

Command A and Command R specifications
Command ACommand R
ProviderCohereCohere
Noometry Index36.531.4
Released2025-03-132024-08-30
WeightsOpenOpen
Context window256K128K
Max output8K4K
Input $ / M tokens$2.50$0.15
Output $ / M tokens$10$0.60
Results tracked2429

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Command R leads

Command A: 27.2 (#322), Command R: 29.3 (#306)

Coding benchmarks
BenchmarkCommand ACommand R
LMArena Coding13301169
Aider Polyglot12%—
BigCodeBench Instruct—37.1%
LiveBench Coding—17.9%
BigCodeBench Complete—45.2%

Agentic & Tool Use Not comparable

Command A: 35.9 (#40), Command R: —

Agentic & Tool Use benchmarks
BenchmarkCommand ACommand R
Berkeley Function Calling Leaderboard57.1%—

Reasoning Command A leads

Command A: 18.3 (#283), Command R: 13.8 (#331)

Reasoning benchmarks
BenchmarkCommand ACommand R
LMArena Hard Prompts13261164
DTBench61.3%46.4%
LMCA10.3%9.2%
Kagi LLM Benchmark28.8%—
LiveBench Reasoning—21.9%
LiveBench Data Analysis—33.3%
LiveBench—27.5%

Math Command A leads

Command A: 36.2 (#171), Command R: 28.0 (#246)

Math benchmarks
BenchmarkCommand ACommand R
LMArena Math13001155
LiveBench Math—19.4%

Knowledge Command A leads

Command A: 37.1 (#159), Command R: 31.0 (#221)

Knowledge benchmarks
BenchmarkCommand ACommand R
LMArena Expert12951138
Vectara Hallucination Rate9.3%—
MMLU—65.2%

Multilingual Command A leads

Command A: 45.3 (#170), Command R: 35.7 (#245)

Multilingual benchmarks
BenchmarkCommand ACommand R
LMArena Non-English13131174
LMArena Chinese13271182
LMArena French13511162
LMArena German13411176
LMArena Japanese12851143
LMArena Korean12851163
LMArena Russian13141174
LMArena Spanish13471151

Instruction Following Command A leads

Command A: 69.1 (#177), Command R: 58.1 (#261)

Instruction Following benchmarks
BenchmarkCommand ACommand R
LMArena Instruction Following13091167
LiveBench Instruction Following—55.6%

Long Context Command A leads

Command A: 40.6 (#151), Command R: 36.3 (#231)

Long Context benchmarks
BenchmarkCommand ACommand R
LMArena Longer Query13341198

Writing & Preference Command A leads

Command A: 47.6 (#208), Command R: 38.2 (#254)

Writing & Preference benchmarks
BenchmarkCommand ACommand R
LMArena Text13311187
LMArena Creative Writing13191170
LMArena Multi-Turn13391163
EQ-Bench Creative Writing1145—
LiveBench Language—16.7%

Frequently asked questions

Is Command A better than Command R?

Command A is the stronger model overall, scoring 36.5 to 31.4 on the Noometry Index. Command R costs 17× less per token, which makes it the better buy when Command A's lead doesn't matter for your workload.

Which is cheaper, Command A or Command R?

Command R is cheaper. It lists at $0.15 per million input tokens and $0.60 per million output tokens; Command A lists at $2.50 and $10.

Is Command A or Command R better for coding?

Command R scores higher on coding benchmarks: 29.3 versus 27.2 in the Noometry coding category.

Which has the bigger context window?

Command A does, with 256K tokens against 128K.

How many benchmarks do Command A and Command R share?

19 benchmarks have published results for both models. Command A has 24 scored results on Noometry and Command R has 29.

Related comparisons

Go deeper