Model comparison

Command R+ vs Mistral Medium 3.1

Command R+ and Mistral Medium 3.1 score almost the same on the Noometry Index (32.4 vs 31.9), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Command R+ Cohere

32.4

Rank #257 Confirmed

Mistral Medium 3.1 Mistral AI

31.9

Rank #266 Reported

Summary

  • The widest gap is in writing & preference, where Mistral Medium 3.1 leads 55.5 to 43.5.
  • Mistral Medium 3.1 is cheaper at $0.40 / $2 per million input/output tokens, against $2.50 / $10 for Command R+.
  • Mistral Medium 3.1 accepts more context: 131K tokens versus 128K.
  • Command R+ has downloadable open weights; the other is API-only.

Side by side

Command R+ and Mistral Medium 3.1 specifications
Command R+Mistral Medium 3.1
ProviderCohereMistral AI
Noometry Index32.431.9
Released2024-08-30—
WeightsOpenProprietary
Context window128K131K
Max output4K105K
Input $ / M tokens$2.50$0.40
Output $ / M tokens$10$2
Results tracked343

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Command R+: 29.1 (#309), Mistral Medium 3.1: —

Coding benchmarks
BenchmarkCommand R+Mistral Medium 3.1
BigCodeBench Instruct33.8%—
LiveBench Coding19.1%—
LMArena Coding1187—
BigCodeBench Complete41.9%—
HumanEval+56.7%—
MBPP+63.5%—

Reasoning Mistral Medium 3.1 leads

Command R+: 9.2 (#344), Mistral Medium 3.1: 10.6 (#341)

Reasoning benchmarks
BenchmarkCommand R+Mistral Medium 3.1
SimpleBench17.4%—
NYT Connections (extended)—6.5%
Thematic Generalization—20.3%
LiveBench Reasoning24.8%—
LMArena Hard Prompts1186—
DTBench54.9%—
LiveBench Data Analysis38.1%—
LMCA5%—
Epoch Capabilities Index119.34—
LiveBench31.8%—

Math Not comparable

Command R+: 28.9 (#242), Mistral Medium 3.1: —

Math benchmarks
BenchmarkCommand R+Mistral Medium 3.1
LiveBench Math21.3%—
LMArena Math1188—

Knowledge Not comparable

Command R+: 36.4 (#169), Mistral Medium 3.1: —

Knowledge benchmarks
BenchmarkCommand R+Mistral Medium 3.1
Vectara Hallucination Rate6.9%—
LMArena Expert1174—
MMLU69.4%—

Multilingual Not comparable

Command R+: 38.6 (#227), Mistral Medium 3.1: —

Multilingual benchmarks
BenchmarkCommand R+Mistral Medium 3.1
LMArena Non-English1216—
LMArena Chinese1226—
LMArena French1209—
LMArena German1216—
LMArena Japanese1166—
LMArena Korean1138—
LMArena Russian1227—
LMArena Spanish1189—

Instruction Following Not comparable

Command R+: 60.0 (#254), Mistral Medium 3.1: —

Instruction Following benchmarks
BenchmarkCommand R+Mistral Medium 3.1
LiveBench Instruction Following57.6%—
LMArena Instruction Following1197—

Long Context Not comparable

Command R+: 37.3 (#219), Mistral Medium 3.1: —

Long Context benchmarks
BenchmarkCommand R+Mistral Medium 3.1
LMArena Longer Query1230—

Writing & Preference Mistral Medium 3.1 leads

Command R+: 43.5 (#228), Mistral Medium 3.1: 55.5 (#145)

Writing & Preference benchmarks
BenchmarkCommand R+Mistral Medium 3.1
LMArena Text1229—
LMArena Creative Writing1235—
EQ-Bench Creative Writing—1476
LMArena Multi-Turn1213—
LiveBench Language29.7%—

Frequently asked questions

Is Command R+ better than Mistral Medium 3.1?

Command R+ and Mistral Medium 3.1 score almost the same on the Noometry Index (32.4 vs 31.9), so choose on price, context window or the category you care about most.

Which is cheaper, Command R+ or Mistral Medium 3.1?

Mistral Medium 3.1 is cheaper. It lists at $0.40 per million input tokens and $2 per million output tokens; Command R+ lists at $2.50 and $10.

Which has the bigger context window?

Mistral Medium 3.1 does, with 131K tokens against 128K.

How many benchmarks do Command R+ and Mistral Medium 3.1 share?

0 benchmarks have published results for both models. Command R+ has 34 scored results on Noometry and Mistral Medium 3.1 has 3.

Related comparisons

Go deeper