Model comparison

Command A vs Gemma 2 27B

Command A is the stronger model overall, scoring 36.5 to 29.4 on the Noometry Index. Gemma 2 27B costs 6.7× less per token, which makes it the better buy when Command A's lead doesn't matter for your workload.

Last verified . 19 shared benchmarks.

Command A Cohere

36.5

Rank #215 Confirmed

Gemma 2 27B Google

29.4

Rank #312 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Command A scores higher in 7 categories and Gemma 2 27B in 1 category; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Command A leads 36.2 to 10.7.
  • The biggest single-benchmark swing is DTBench: 61.3% for Command A and 48% for Gemma 2 27B.
  • Gemma 2 27B is cheaper at $0.65 / $0.65 per million input/output tokens, against $2.50 / $10 for Command A.
  • Command A accepts more context: 256K tokens versus 8K.

Side by side

Command A and Gemma 2 27B specifications
Command AGemma 2 27B
ProviderCohereGoogle
Noometry Index36.529.4
Released2025-03-132024-06-24
WeightsOpenOpen
Context window256K8K
Max output8K2K
Input $ / M tokens$2.50$0.65
Output $ / M tokens$10$0.65
Results tracked2434

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemma 2 27B leads

Command A: 27.2 (#322), Gemma 2 27B: 34.1 (#246)

Coding benchmarks
BenchmarkCommand AGemma 2 27B
LMArena Coding13301211
Aider Polyglot12%—
BigCodeBench Instruct—42.8%
LiveBench Coding—36%
BigCodeBench Complete—52.5%

Agentic & Tool Use Not comparable

Command A: 35.9 (#40), Gemma 2 27B: —

Agentic & Tool Use benchmarks
BenchmarkCommand AGemma 2 27B
Berkeley Function Calling Leaderboard57.1%—

Reasoning Command A leads

Command A: 18.3 (#283), Gemma 2 27B: 15.3 (#315)

Reasoning benchmarks
BenchmarkCommand AGemma 2 27B
LMArena Hard Prompts13261198
DTBench61.3%48%
LMCA10.3%7.1%
Kagi LLM Benchmark28.8%—
LiveBench Reasoning—28.1%
LiveBench Data Analysis—47.9%
Epoch Capabilities Index—122.08
LiveBench—38.2%

Math Command A leads

Command A: 36.2 (#171), Gemma 2 27B: 10.7 (#311)

Math benchmarks
BenchmarkCommand AGemma 2 27B
LMArena Math13001212
OTIS Mock AIME 2024-2025—1.4%
LiveBench Math—26.5%
MATH Level 5—27.9%

Knowledge Command A leads

Command A: 37.1 (#159), Gemma 2 27B: 19.0 (#280)

Knowledge benchmarks
BenchmarkCommand AGemma 2 27B
LMArena Expert12951172
GPQA Diamond—36.5%
Confabulations—27.1%
Vectara Hallucination Rate9.3%—
MMLU—75.7%

Multilingual Command A leads

Command A: 45.3 (#170), Gemma 2 27B: 38.6 (#226)

Multilingual benchmarks
BenchmarkCommand AGemma 2 27B
LMArena Non-English13131217
LMArena Chinese13271221
LMArena French13511247
LMArena German13411209
LMArena Japanese12851175
LMArena Korean12851174
LMArena Russian13141234
LMArena Spanish13471228

Instruction Following Command A leads

Command A: 69.1 (#177), Gemma 2 27B: 60.5 (#249)

Instruction Following benchmarks
BenchmarkCommand AGemma 2 27B
LMArena Instruction Following13091206
LiveBench Instruction Following—58.1%

Long Context Command A leads

Command A: 40.6 (#151), Gemma 2 27B: 37.3 (#218)

Long Context benchmarks
BenchmarkCommand AGemma 2 27B
LMArena Longer Query13341231

Writing & Preference Command A leads

Command A: 47.6 (#208), Gemma 2 27B: 44.2 (#225)

Writing & Preference benchmarks
BenchmarkCommand AGemma 2 27B
LMArena Text13311231
LMArena Creative Writing13191241
LMArena Multi-Turn13391224
EQ-Bench Creative Writing1145—
LiveBench Language—32.6%

Frequently asked questions

Is Command A better than Gemma 2 27B?

Command A is the stronger model overall, scoring 36.5 to 29.4 on the Noometry Index. Gemma 2 27B costs 6.7× less per token, which makes it the better buy when Command A's lead doesn't matter for your workload.

Which is cheaper, Command A or Gemma 2 27B?

Gemma 2 27B is cheaper. It lists at $0.65 per million input tokens and $0.65 per million output tokens; Command A lists at $2.50 and $10.

Is Command A or Gemma 2 27B better for coding?

Gemma 2 27B scores higher on coding benchmarks: 34.1 versus 27.2 in the Noometry coding category.

Which has the bigger context window?

Command A does, with 256K tokens against 8K.

How many benchmarks do Command A and Gemma 2 27B share?

19 benchmarks have published results for both models. Command A has 24 scored results on Noometry and Gemma 2 27B has 34.

Related comparisons

Go deeper