Cohere, open weights
Command R+
Command R+ by Cohere ranks 257th of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.4. Its strongest category is knowledge, where it ranks 169th. API pricing starts at $2.50 per million input tokens and $10 per million output tokens, with a 128K-token context window.
Last verified
Specifications
- Noometry rank
- #257 of 354
- Index score
- 32.4
- Evidence
- Confirmed 34 results
- Provider
Cohere
- Released
- August 30, 2024
- Weights
- Open weights
- Reasoning
- No
- Context window
- 128K
- Max output
- 4K
- Input price
- $2.50 / M
- Output price
- $10 / M
- Blended price
- $4.38 / M
- Output speed
- Not measured
- Value
- #195 of 219
- Knowledge cutoff
- June 2024
- Input
- text
Category scores
Each category score combines every public result we have in that category.
- Coding 29.1
- Reasoning 9.2
- Math 28.9
- Knowledge 36.4
- Multilingual 38.6
- Instruction Following 60.0
- Long Context 37.3
- Writing & Preference 43.5
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 29.1 | #309 | 4 |
| Reasoning | 9.2 | #344 | 6 |
| Math | 28.9 | #242 | 2 |
| Knowledge | 36.4 | #169 | 2 |
| Multilingual | 38.6 | #227 | 1 |
| Instruction Following | 60.0 | #254 | 2 |
| Long Context | 37.3 | #219 | 1 |
| Writing & Preference | 43.5 | #228 | 4 |
Strengths and weaknesses
Categories where Command R+ places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Knowledge | 36.4 | −0.9 | #169 of 314, top 54% |
| Writing & Preference | 43.5 | −10.3 | #228 of 312, top 74% |
| Long Context | 37.3 | −3.6 | #219 of 296, top 74% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Reasoning | 9.2 | −14.4 | #344 of 350, top 99% |
| Coding | 29.1 | −9.6 | #309 of 340, top 91% |
| Instruction Following | 60.0 | −11.3 | #254 of 305, top 84% |
Closest competitors
The models ranked just above and below Command R+. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Qwen1.5-14B | #253 | 32.7 | — | — | Compare |
| Olmo 2 0325 32b Instruct | #254 | 32.7 | — | — | Compare |
| gpt-oss-20b | #255 | 32.5 | $0.036 | 96 | Compare |
| Laguna M.1 | #256 | 32.5 | — | — | Compare |
| Granite 3.1 8b Instruct | #258 | 32.4 | — | — | Compare |
| Pixtral Large | #259 | 32.2 | $3 | — | Compare |
| Falcon-180B | #260 | 32.2 | — | — | Compare |
| Gemini 1.5 Pro (May 2024) | #261 | 32.1 | — | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| BigCodeBench Instruct | 33.8% | #49 of 64, top 77% | BigCodeBench | 2024-04-04 | |
| LiveBench Coding | 19.1% | #36 of 39, top 93% | Epoch AI | ||
| LMArena Coding | 1187 | #245 of 294, top 84% | LMArena | 2026-10-08 | |
| LMArena Coding | 1172 | LMArena | 2026-10-08 | ||
| BigCodeBench Complete | 41.9% | #49 of 66, top 75% | BigCodeBench | 2024-04-04 | |
| HumanEval+ | 56.7% | #31 of 45, top 69% | EvalPlus | ||
| MBPP+ | 63.5% | #21 of 38, top 56% | EvalPlus |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| SimpleBench | 17.4% | #76 of 77, top 99% | Epoch AI | ||
| LiveBench Reasoning | 24.8% | #35 of 39, top 90% | Epoch AI | ||
| LMArena Hard Prompts | 1186 | #244 of 297, top 83% | LMArena | 2026-10-08 | |
| LMArena Hard Prompts | 1166 | LMArena | 2026-10-08 | ||
| DTBench | 54.9% | #123 of 151, top 82% | Epoch AI | ||
| LiveBench Data Analysis | 38.1% | #31 of 39, top 80% | Epoch AI | ||
| LMCA | 5% | #123 of 125, top 99% | Epoch AI | ||
| Epoch Capabilities Index | 119.34 | #171 of 213, top 81% | Epoch AI | 2024-08-30 | |
| LiveBench | 31.8% | #33 of 39, top 85% | Epoch AI |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LiveBench Math | 21.3% | #34 of 39, top 88% | Epoch AI | ||
| LMArena Math | 1164 | LMArena | 2026-10-08 | ||
| LMArena Math | 1188 | #236 of 285, top 83% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Vectara Hallucination Rate (lower is better) | 6.9% | #26 of 96, top 28% | Vectara Hallucination Leaderboard | ||
| LMArena Expert | 1174 | #227 of 273, top 84% | LMArena | 2026-10-08 | |
| LMArena Expert | 1153 | LMArena | 2026-10-08 | ||
| MMLU | 69.4% | #47 of 81, top 59% | Epoch AI |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1190 | LMArena | 2026-10-08 | ||
| LMArena Non-English | 1216 | #227 of 297, top 77% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1226 | #220 of 285, top 78% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1191 | LMArena | 2026-10-08 | ||
| LMArena French | 1209 | #191 of 223, top 86% | LMArena | 2026-10-08 | |
| LMArena German | 1195 | LMArena | 2026-10-08 | ||
| LMArena German | 1216 | #187 of 231, top 81% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1166 | #170 of 211, top 81% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1155 | LMArena | 2026-10-08 | ||
| LMArena Korean | 1138 | #184 of 213, top 87% | LMArena | 2026-10-08 | |
| LMArena Russian | 1227 | #223 of 283, top 79% | LMArena | 2026-10-08 | |
| LMArena Russian | 1206 | LMArena | 2026-10-08 | ||
| LMArena Spanish | 1189 | #197 of 226, top 88% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LiveBench Instruction Following | 57.6% | #31 of 39, top 80% | Epoch AI | ||
| LMArena Instruction Following | 1197 | #238 of 298, top 80% | LMArena | 2026-10-08 | |
| LMArena Instruction Following | 1180 | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1230 | #230 of 291, top 80% | LMArena | 2026-10-08 | |
| LMArena Longer Query | 1202 | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1204 | LMArena | 2026-10-08 | ||
| LMArena Text | 1229 | #231 of 297, top 78% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1235 | #218 of 295, top 74% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1201 | LMArena | 2026-10-08 | ||
| LMArena Multi-Turn | 1192 | LMArena | 2026-10-08 | ||
| LMArena Multi-Turn | 1213 | #235 of 295, top 80% | LMArena | 2026-10-08 | |
| LiveBench Language | 29.7% | #27 of 39, top 70% | Epoch AI |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| cohere | $2.50 | $10 | — | 2026-10-10 |
| openrouter | $2.50 | $10 | — | 2026-10-10 |
Compare Command R+
- Command R+ vs Laguna M.1
- Command R+ vs Granite 3.1 8b Instruct
- Command R+ vs gpt-oss-20b
- Command R+ vs Pixtral Large
- Command R+ vs Olmo 2 0325 32b Instruct
- Command R+ vs Falcon-180B
- Command R+ vs GPT-6 Astra
- Command R+ vs Claude Fable 5.1
- Command R+ vs Gemini 3.8 Flash
- Command R+ vs Kimi K3
- Command R+ vs Grok 4.6
- Command R+ vs Qwen3.8 Max
- Command R+ vs GLM-5.3
- Command R+ vs Muse Spark 1.3
Other Cohere models
Frequently asked questions
How good is Command R+?
Command R+ by Cohere ranks 257th of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.4. Its strongest category is knowledge, where it ranks 169th. API pricing starts at $2.50 per million input tokens and $10 per million output tokens, with a 128K-token context window.
How much does Command R+ cost?
Command R+ costs $2.50 per million input tokens and $10 per million output tokens on Cohere's own API.
What is Command R+'s context window?
Command R+ accepts up to 128K tokens of input and can write up to 4K tokens in one response.
Is Command R+ open source?
Yes. Command R+'s weights are downloadable; check the license for commercial terms.
What are Command R+'s strengths and weaknesses?
Relative to other ranked models, Command R+ places best in knowledge, writing & preference, long context and lowest in reasoning, coding, instruction following.
What is Command R+ best at?
Its best category is knowledge, where it ranks 169th on Noometry.