Model comparison

Command A vs Grok 4.1 Fast

Grok 4.1 Fast is the stronger model overall, scoring 41.4 to 36.5 on the Noometry Index.

Last verified . 21 shared benchmarks.

Command A Cohere

36.5

Rank #215 Confirmed

Grok 4.1 Fast xAI

41.4

Rank #136 Confirmed

Summary

  • They share 21 benchmarks with published results for both. Command A scores higher in 2 categories and Grok 4.1 Fast in 7 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Grok 4.1 Fast leads 43.4 to 18.3.
  • The biggest single-benchmark swing is DTBench: 61.3% for Command A and 87.7% for Grok 4.1 Fast.
  • Grok 4.1 Fast is cheaper at $0.20 / $0.50 per million input/output tokens, against $2.50 / $10 for Command A.
  • Command A accepts more context: 256K tokens versus 128K.
  • Command A has downloadable open weights; the other is API-only.

Side by side

Command A and Grok 4.1 Fast specifications
Command AGrok 4.1 Fast
ProviderCoherexAI
Noometry Index36.541.4
Released2025-03-132025-06-27
WeightsOpenProprietary
Context window256K128K
Max output8K30K
Input $ / M tokens$2.50$0.20
Output $ / M tokens$10$0.50
Results tracked2432

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Grok 4.1 Fast leads

Command A: 27.2 (#322), Grok 4.1 Fast: 34.1 (#245)

Coding benchmarks
BenchmarkCommand AGrok 4.1 Fast
LMArena Coding13301411
Aider Polyglot12%—
LMArena WebDev—1242
ALE-Bench—394.93

Agentic & Tool Use Too close to call

Command A: 35.9 (#40), Grok 4.1 Fast: 36.3 (#39)

Agentic & Tool Use benchmarks
BenchmarkCommand AGrok 4.1 Fast
Berkeley Function Calling Leaderboard57.1%69.6%
τ²-bench Banking—13.1%
LMArena Search—1171
Vending-Bench 2—1,107

Reasoning Grok 4.1 Fast leads

Command A: 18.3 (#283), Grok 4.1 Fast: 43.4 (#49)

Reasoning benchmarks
BenchmarkCommand AGrok 4.1 Fast
LMArena Hard Prompts13261407
DTBench61.3%87.7%
SimpleBench—56%
Kagi LLM Benchmark28.8%—
NYT Connections (extended)—87.4%
LMCA10.3%—
ForecastBench—61

Math Command A leads

Command A: 36.2 (#171), Grok 4.1 Fast: 31.9 (#221)

Math benchmarks
BenchmarkCommand AGrok 4.1 Fast
LMArena Math13001408
MathArena Final-Answer Competitions—60.9%
ProofBench—4%

Knowledge Command A leads

Command A: 37.1 (#159), Grok 4.1 Fast: 33.1 (#207)

Knowledge benchmarks
BenchmarkCommand AGrok 4.1 Fast
Vectara Hallucination Rate9.3%17.8%
LMArena Expert12951399

Multimodal Not comparable

Command A: —, Grok 4.1 Fast: 37.0 (#76)

Multimodal benchmarks
BenchmarkCommand AGrok 4.1 Fast
LMArena Vision—1201

Multilingual Grok 4.1 Fast leads

Command A: 45.3 (#170), Grok 4.1 Fast: 51.0 (#114)

Multilingual benchmarks
BenchmarkCommand AGrok 4.1 Fast
LMArena Non-English13131391
LMArena Chinese13271441
LMArena French13511415
LMArena German13411404
LMArena Japanese12851349
LMArena Korean12851361
LMArena Russian13141387
LMArena Spanish13471413

Instruction Following Grok 4.1 Fast leads

Command A: 69.1 (#177), Grok 4.1 Fast: 72.7 (#133)

Instruction Following benchmarks
BenchmarkCommand AGrok 4.1 Fast
LMArena Instruction Following13091376

Long Context Grok 4.1 Fast leads

Command A: 40.6 (#151), Grok 4.1 Fast: 42.4 (#126)

Long Context benchmarks
BenchmarkCommand AGrok 4.1 Fast
LMArena Longer Query13341390

Writing & Preference Grok 4.1 Fast leads

Command A: 47.6 (#208), Grok 4.1 Fast: 57.2 (#131)

Writing & Preference benchmarks
BenchmarkCommand AGrok 4.1 Fast
LMArena Text13311408
LMArena Creative Writing13191394
EQ-Bench Creative Writing11451327
LMArena Multi-Turn13391389

Frequently asked questions

Is Command A better than Grok 4.1 Fast?

Grok 4.1 Fast is the stronger model overall, scoring 41.4 to 36.5 on the Noometry Index.

Which is cheaper, Command A or Grok 4.1 Fast?

Grok 4.1 Fast is cheaper. It lists at $0.20 per million input tokens and $0.50 per million output tokens; Command A lists at $2.50 and $10.

Is Command A or Grok 4.1 Fast better for coding?

Grok 4.1 Fast scores higher on coding benchmarks: 34.1 versus 27.2 in the Noometry coding category.

Which has the bigger context window?

Command A does, with 256K tokens against 128K.

How many benchmarks do Command A and Grok 4.1 Fast share?

21 benchmarks have published results for both models. Command A has 24 scored results on Noometry and Grok 4.1 Fast has 32.

Related comparisons

Go deeper