Cohere, open weights

Command A

Command A by Cohere ranks 215th of 354 ranked models on the Noometry Index as of October 2026, with a score of 36.5. Its strongest category is agentic & tool use, where it ranks 40th. API pricing starts at $2.50 per million input tokens and $10 per million output tokens, with a 256K-token context window.

Last verified

Specifications

Noometry rank
#215 of 354
Index score
36.5
Evidence
Confirmed 24 results
Provider
Cohere
Released
March 13, 2025
Weights
Open weights
Reasoning
No
Context window
256K
Max output
8K
Input price
$2.50 / M
Output price
$10 / M
Blended price
$4.38 / M
Output speed
28 tokens/s Kagi
Value
#192 of 219
Knowledge cutoff
June 2024
Input
text

Category scores

Each category score combines every public result we have in that category.

Command A category scores
  1. Coding 27.2
  2. Agentic & Tool Use 35.9
  3. Reasoning 18.3
  4. Math 36.2
  5. Knowledge 37.1
  6. Multilingual 45.3
  7. Instruction Following 69.1
  8. Long Context 40.6
  9. Writing & Preference 47.6
Command A category ranks
CategoryScoreRankResults
Coding27.2#3222
Agentic & Tool Use35.9#401
Reasoning18.3#2834
Math36.2#1711
Knowledge37.1#1592
Multilingual45.3#1701
Instruction Following69.1#1771
Long Context40.6#1511
Writing & Preference47.6#2084

Strengths and weaknesses

Categories where Command A places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Command A: strongest categories
CategoryScorevs medianRank
Agentic & Tool Use35.9+5.6#40 of 154, top 26%
Knowledge37.1−0.2#159 of 314, top 51%
Long Context40.6−0.3#151 of 296, top 52%

Weakest categories

Command A: weakest categories
CategoryScorevs medianRank
Coding27.2−11.5#322 of 340, top 95%
Reasoning18.3−5.3#283 of 350, top 81%
Writing & Preference47.6−6.2#208 of 312, top 67%

Closest competitors

The models ranked just above and below Command A. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Command A
ModelRankScoreBlended $/MSpeed
Gemini 2.5 Flash-Lite#21137.0$0.18172Compare
o3-mini#21236.7$1.93—Compare
Llama 3.1 Nemotron Ultra 253b v1#21336.7——Compare
Granite 4.0 H Small#21436.5——Compare
Grok Build 0.1#21636.4$1.25—Compare
gpt-oss-120b#21736.3$0.070355Compare
Mistral Medium#21836.3$368Compare
GPT-4.1#21935.9$3.50116Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Command A Coding benchmark results
BenchmarkScorePositionSettingSourceDate
Aider Polyglot12%#40 of 44, top 91%Epoch AI
LMArena Coding1330#179 of 294, top 61%LMArena2026-10-08

Agentic & Tool Use

Command A Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
Berkeley Function Calling Leaderboard46.5%fcBerkeley Function Calling Leaderboard
Berkeley Function Calling Leaderboard57.1%#10 of 49, top 21%fcBerkeley Function Calling Leaderboard

Reasoning

Command A Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Kagi LLM Benchmark28.8%#93 of 99, top 94%Kagi LLM Benchmark
LMArena Hard Prompts1326#176 of 297, top 60%LMArena2026-10-08
DTBench61.3%#114 of 151, top 76%Epoch AI
LMCA10.3%#111 of 125, top 89%Epoch AI

Math

Command A Math benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Math1300#183 of 285, top 65%LMArena2026-10-08

Knowledge

Command A Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
Vectara Hallucination Rate (lower is better)9.3%#43 of 96, top 45%Vectara Hallucination Leaderboard
LMArena Expert1295#178 of 273, top 66%LMArena2026-10-08

Multilingual

Command A Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1313#170 of 297, top 58%LMArena2026-10-08
LMArena Chinese1327#179 of 285, top 63%LMArena2026-10-08
LMArena French1351#144 of 223, top 65%LMArena2026-10-08
LMArena German1341#135 of 231, top 59%LMArena2026-10-08
LMArena Japanese1285#130 of 211, top 62%LMArena2026-10-08
LMArena Korean1285#135 of 213, top 64%LMArena2026-10-08
LMArena Russian1314#169 of 283, top 60%LMArena2026-10-08
LMArena Spanish1347#146 of 226, top 65%LMArena2026-10-08

Instruction Following

Command A Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1309#170 of 298, top 58%LMArena2026-10-08

Long Context

Command A Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1334#160 of 291, top 55%LMArena2026-10-08

Writing & Preference

Command A Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1331#172 of 297, top 58%LMArena2026-10-08
LMArena Creative Writing1319#153 of 295, top 52%LMArena2026-10-08
EQ-Bench Creative Writing1145#89 of 115, top 78%EQ-Bench
LMArena Multi-Turn1339#165 of 295, top 56%LMArena2026-10-08

API pricing by provider

Command A API prices
RouteInput $/MOutput $/MCached input $/MChecked
azure$2.50$10—2026-10-10
cohere$2.50$10—2026-10-10
openrouter$2.50$10—2026-10-10

Compare Command A

Other Cohere models

Frequently asked questions

How good is Command A?

Command A by Cohere ranks 215th of 354 ranked models on the Noometry Index as of October 2026, with a score of 36.5. Its strongest category is agentic & tool use, where it ranks 40th. API pricing starts at $2.50 per million input tokens and $10 per million output tokens, with a 256K-token context window.

How much does Command A cost?

Command A costs $2.50 per million input tokens and $10 per million output tokens on Cohere's own API.

What is Command A's context window?

Command A accepts up to 256K tokens of input and can write up to 8K tokens in one response.

Is Command A open source?

Yes. Command A's weights are downloadable from Hugging Face (CohereForAI/c4ai-command-a-03-2025); check the license for commercial terms.

How fast is Command A?

Command A generated about 28 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.

What are Command A's strengths and weaknesses?

Relative to other ranked models, Command A places best in agentic & tool use, knowledge, long context and lowest in coding, reasoning, writing & preference.

What is Command A best at?

Its best category is agentic & tool use, where it ranks 40th on Noometry.