Model comparison

Command A vs Gemma 3 27B

Command A is the stronger model overall, scoring 36.5 to 30.8 on the Noometry Index. Gemma 3 27B costs 44× less per token, which makes it the better buy when Command A's lead doesn't matter for your workload.

Last verified . 24 shared benchmarks.

Command A Cohere

36.5

Rank #215 Confirmed

Gemma 3 27B Google

30.8

Rank #284 Confirmed

Summary

  • They share 24 benchmarks with published results for both. Command A scores higher in 6 categories and Gemma 3 27B in 3 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in long context, where Command A leads 40.6 to 27.6.
  • The biggest single-benchmark swing is Berkeley Function Calling Leaderboard: 57.1% for Command A and 29.5% for Gemma 3 27B.
  • Gemma 3 27B is cheaper at $0.08 / $0.16 per million input/output tokens, against $2.50 / $10 for Command A.
  • Command A accepts more context: 256K tokens versus 131K.

Side by side

Command A and Gemma 3 27B specifications
Command AGemma 3 27B
ProviderCohereGoogle
Noometry Index36.530.8
Released2025-03-132025-03-11
WeightsOpenOpen
Context window256K131K
Max output8K8K
Input $ / M tokens$2.50$0.08
Output $ / M tokens$10$0.16
Results tracked2443

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Command A leads

Command A: 27.2 (#322), Gemma 3 27B: 22.5 (#334)

Coding benchmarks
BenchmarkCommand AGemma 3 27B
Aider Polyglot12%4.9%
LMArena Coding13301322
SciCode—21.2%
LiveBench Coding—39.9%

Agentic & Tool Use Command A leads

Command A: 35.9 (#40), Gemma 3 27B: 25.1 (#110)

Agentic & Tool Use benchmarks
BenchmarkCommand AGemma 3 27B
Berkeley Function Calling Leaderboard57.1%29.5%

Reasoning Command A leads

Command A: 18.3 (#283), Gemma 3 27B: 16.7 (#301)

Reasoning benchmarks
BenchmarkCommand AGemma 3 27B
Kagi LLM Benchmark28.8%40.4%
LMArena Hard Prompts13261340
DTBench61.3%52.5%
LMCA10.3%12.3%
CritPt—0%
Chess Puzzles—0%
LiveBench Reasoning—43.8%
LiveBench Data Analysis—51.5%
Epoch Capabilities Index—130.04
LiveBench—50%

Math Command A leads

Command A: 36.2 (#171), Gemma 3 27B: 25.9 (#265)

Math benchmarks
BenchmarkCommand AGemma 3 27B
LMArena Math13001312
OTIS Mock AIME 2024-2025—22.5%
LiveBench Math—55.4%
MATH Level 5—74%

Knowledge Command A leads

Command A: 37.1 (#159), Gemma 3 27B: 25.5 (#261)

Knowledge benchmarks
BenchmarkCommand AGemma 3 27B
Vectara Hallucination Rate9.3%7.4%
LMArena Expert12951304
GPQA Diamond—47.7%
Confabulations—40.3%

Multimodal Not comparable

Command A: —, Gemma 3 27B: 32.6 (#100)

Multimodal benchmarks
BenchmarkCommand AGemma 3 27B
LMArena Vision—1164
GeoBench—52%

Multilingual Gemma 3 27B leads

Command A: 45.3 (#170), Gemma 3 27B: 46.9 (#155)

Multilingual benchmarks
BenchmarkCommand AGemma 3 27B
LMArena Non-English13131334
LMArena Chinese13271346
LMArena French13511368
LMArena German13411362
LMArena Japanese12851287
LMArena Korean12851308
LMArena Russian13141349
LMArena Spanish13471349

Instruction Following Gemma 3 27B leads

Command A: 69.1 (#177), Gemma 3 27B: 70.6 (#160)

Instruction Following benchmarks
BenchmarkCommand AGemma 3 27B
LMArena Instruction Following13091321
LiveBench Instruction Following—74.9%

Long Context Command A leads

Command A: 40.6 (#151), Gemma 3 27B: 27.6 (#293)

Long Context benchmarks
BenchmarkCommand AGemma 3 27B
LMArena Longer Query13341333
Fiction.LiveBench—33.3%

Writing & Preference Gemma 3 27B leads

Command A: 47.6 (#208), Gemma 3 27B: 52.5 (#168)

Writing & Preference benchmarks
BenchmarkCommand AGemma 3 27B
LMArena Text13311358
LMArena Creative Writing13191346
EQ-Bench Creative Writing11451266
LMArena Multi-Turn13391345
Short-Story Creative Writing—79.9%
LiveBench Language—34.6%

Frequently asked questions

Is Command A better than Gemma 3 27B?

Command A is the stronger model overall, scoring 36.5 to 30.8 on the Noometry Index. Gemma 3 27B costs 44× less per token, which makes it the better buy when Command A's lead doesn't matter for your workload.

Which is cheaper, Command A or Gemma 3 27B?

Gemma 3 27B is cheaper. It lists at $0.08 per million input tokens and $0.16 per million output tokens; Command A lists at $2.50 and $10.

Is Command A or Gemma 3 27B better for coding?

Command A scores higher on coding benchmarks: 27.2 versus 22.5 in the Noometry coding category.

Which has the bigger context window?

Command A does, with 256K tokens against 131K.

How many benchmarks do Command A and Gemma 3 27B share?

24 benchmarks have published results for both models. Command A has 24 scored results on Noometry and Gemma 3 27B has 43.

Related comparisons

Go deeper