Model comparison

Command A vs Gemini 3 Flash Preview

Gemini 3 Flash Preview is the stronger model overall, scoring 52.3 to 36.5 on the Noometry Index.

Last verified . 20 shared benchmarks.

Command A Cohere

36.5

Rank #215 Confirmed

Gemini 3 Flash Preview Google

52.3

Rank #40 Confirmed

Summary

  • They share 20 benchmarks with published results for both. Command A scores higher in 0 categories and Gemini 3 Flash Preview in 9 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Gemini 3 Flash Preview leads 49.2 to 18.3.
  • The biggest single-benchmark swing is LMCA: 10.3% for Command A and 43.1% for Gemini 3 Flash Preview.
  • Gemini 3 Flash Preview is cheaper at $0.50 / $3 per million input/output tokens, against $2.50 / $10 for Command A.
  • Gemini 3 Flash Preview accepts more context: 1.05M tokens versus 256K.
  • Command A has downloadable open weights; the other is API-only.

Side by side

Command A and Gemini 3 Flash Preview specifications
Command AGemini 3 Flash Preview
ProviderCohereGoogle
Noometry Index36.552.3
Released2025-03-132025-12-17
WeightsOpenProprietary
Context window256K1.05M
Max output8K66K
Input $ / M tokens$2.50$0.50
Output $ / M tokens$10$3
Results tracked2459

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 3 Flash Preview leads

Command A: 27.2 (#322), Gemini 3 Flash Preview: 50.9 (#42)

Coding benchmarks
BenchmarkCommand AGemini 3 Flash Preview
LMArena Coding13301460
SWE-bench Verified—75.4%
SWE-bench Verified (bash only)—75.8%
Aider Polyglot12%—
LMArena WebDev—1439
SWE-bench Multilingual—72.7%
GSO—9.8%
WeirdML—61.6%
ALE-Bench—1,367

Agentic & Tool Use Gemini 3 Flash Preview leads

Command A: 35.9 (#40), Gemini 3 Flash Preview: 38.7 (#29)

Agentic & Tool Use benchmarks
BenchmarkCommand AGemini 3 Flash Preview
Terminal-Bench—64.3%
Berkeley Function Calling Leaderboard57.1%—
τ²-bench Airline—82.5%
τ²-bench Banking—27.3%
τ²-bench Retail—76.8%
τ²-bench Telecom—91.2%
DeepResearch Bench—49.8%
BALROG—48.1%
GDP.pdf—10%
LMArena Search—1198
Vending-Bench 2—3,635

Reasoning Gemini 3 Flash Preview leads

Command A: 18.3 (#283), Gemini 3 Flash Preview: 49.2 (#37)

Reasoning benchmarks
BenchmarkCommand AGemini 3 Flash Preview
LMArena Hard Prompts13261465
DTBench61.3%89.1%
LMCA10.3%43.1%
ARC-AGI-2—33.6%
SimpleBench—61.1%
Kagi LLM Benchmark28.8%—
NYT Connections (extended)—83.1%
ARC-AGI-1—84.7%
Chess Puzzles—40%
Mystery Game Puzzles—26%
Epoch Capabilities Index—151.8
ForecastBench—58.5

Math Gemini 3 Flash Preview leads

Command A: 36.2 (#171), Gemini 3 Flash Preview: 51.7 (#55)

Knowledge Gemini 3 Flash Preview leads

Command A: 37.1 (#159), Gemini 3 Flash Preview: 58.8 (#33)

Knowledge benchmarks
BenchmarkCommand AGemini 3 Flash Preview
Vectara Hallucination Rate9.3%13.5%
LMArena Expert12951462
GPQA Diamond—89.4%
SimpleQA Verified—66.8%

Multimodal Not comparable

Command A: —, Gemini 3 Flash Preview: 45.5 (#16)

Multimodal benchmarks
BenchmarkCommand AGemini 3 Flash Preview
LMArena Vision—1285
GeoBench—88%
VPCT—72.6%
Blueprint-Bench 2—0%
LMArena Document—1413

Multilingual Gemini 3 Flash Preview leads

Command A: 45.3 (#170), Gemini 3 Flash Preview: 55.7 (#27)

Multilingual benchmarks
BenchmarkCommand AGemini 3 Flash Preview
LMArena Non-English13131458
LMArena Chinese13271511
LMArena French13511477
LMArena German13411497
LMArena Japanese12851489
LMArena Korean12851443
LMArena Russian13141480
LMArena Spanish13471469

Instruction Following Gemini 3 Flash Preview leads

Command A: 69.1 (#177), Gemini 3 Flash Preview: 75.7 (#56)

Instruction Following benchmarks
BenchmarkCommand AGemini 3 Flash Preview
LMArena Instruction Following13091437

Long Context Gemini 3 Flash Preview leads

Command A: 40.6 (#151), Gemini 3 Flash Preview: 44.4 (#67)

Long Context benchmarks
BenchmarkCommand AGemini 3 Flash Preview
LMArena Longer Query13341452

Writing & Preference Gemini 3 Flash Preview leads

Command A: 47.6 (#208), Gemini 3 Flash Preview: 65.5 (#45)

Writing & Preference benchmarks
BenchmarkCommand AGemini 3 Flash Preview
LMArena Text13311466
LMArena Creative Writing13191457
LMArena Multi-Turn13391471
EQ-Bench Creative Writing1145—

Frequently asked questions

Is Command A better than Gemini 3 Flash Preview?

Gemini 3 Flash Preview is the stronger model overall, scoring 52.3 to 36.5 on the Noometry Index.

Which is cheaper, Command A or Gemini 3 Flash Preview?

Gemini 3 Flash Preview is cheaper. It lists at $0.50 per million input tokens and $3 per million output tokens; Command A lists at $2.50 and $10.

Is Command A or Gemini 3 Flash Preview better for coding?

Gemini 3 Flash Preview scores higher on coding benchmarks: 50.9 versus 27.2 in the Noometry coding category.

Which has the bigger context window?

Gemini 3 Flash Preview does, with 1.05M tokens against 256K.

How many benchmarks do Command A and Gemini 3 Flash Preview share?

20 benchmarks have published results for both models. Command A has 24 scored results on Noometry and Gemini 3 Flash Preview has 59.

Related comparisons

Go deeper