Model comparison

Command R vs Mistral Small 3.2

Command R and Mistral Small 3.2 score almost the same on the Noometry Index (31.4 vs 31.2), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Mistral Small 3.2 Mistral AI

31.2

Rank #280 Confirmed

Summary

  • The widest gap is in writing & preference, where Mistral Small 3.2 leads 45.0 to 38.2.
  • Mistral Small 3.2 is cheaper at $0.0938 / $0.25 per million input/output tokens, against $0.15 / $0.60 for Command R.
  • Mistral Small 3.2 accepts more context: 256K tokens versus 128K.

Side by side

Command R and Mistral Small 3.2 specifications
Command RMistral Small 3.2
ProviderCohereMistral AI
Noometry Index31.431.2
Released2024-08-302025-06-20
WeightsOpenOpen
Context window128K256K
Max output4K16K
Input $ / M tokens$0.15$0.0938
Output $ / M tokens$0.60$0.25
Results tracked296

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Command R: 29.3 (#306), Mistral Small 3.2: —

Coding benchmarks
BenchmarkCommand RMistral Small 3.2
BigCodeBench Instruct37.1%—
LiveBench Coding17.9%—
LMArena Coding1169—
BigCodeBench Complete45.2%—

Reasoning Mistral Small 3.2 leads

Command R: 13.8 (#331), Mistral Small 3.2: 18.1 (#287)

Reasoning benchmarks
BenchmarkCommand RMistral Small 3.2
Kagi LLM Benchmark—40.4%
Chess Puzzles—1%
LiveBench Reasoning21.9%—
LMArena Hard Prompts1164—
DTBench46.4%—
LiveBench Data Analysis33.3%—
LMCA9.2%—
Epoch Capabilities Index—131.74
LiveBench27.5%—

Math Command R leads

Command R: 28.0 (#246), Mistral Small 3.2: 26.3 (#260)

Math benchmarks
BenchmarkCommand RMistral Small 3.2
OTIS Mock AIME 2024-2025—30.3%
LiveBench Math19.4%—
LMArena Math1155—

Knowledge Command R leads

Command R: 31.0 (#221), Mistral Small 3.2: 26.7 (#256)

Knowledge benchmarks
BenchmarkCommand RMistral Small 3.2
GPQA Diamond—49.1%
LMArena Expert1138—
MMLU65.2%—

Multilingual Not comparable

Command R: 35.7 (#245), Mistral Small 3.2: —

Multilingual benchmarks
BenchmarkCommand RMistral Small 3.2
LMArena Non-English1174—
LMArena Chinese1182—
LMArena French1162—
LMArena German1176—
LMArena Japanese1143—
LMArena Korean1163—
LMArena Russian1174—
LMArena Spanish1151—

Instruction Following Not comparable

Command R: 58.1 (#261), Mistral Small 3.2: —

Instruction Following benchmarks
BenchmarkCommand RMistral Small 3.2
LiveBench Instruction Following55.6%—
LMArena Instruction Following1167—

Long Context Not comparable

Command R: 36.3 (#231), Mistral Small 3.2: —

Long Context benchmarks
BenchmarkCommand RMistral Small 3.2
LMArena Longer Query1198—

Writing & Preference Mistral Small 3.2 leads

Command R: 38.2 (#254), Mistral Small 3.2: 45.0 (#224)

Writing & Preference benchmarks
BenchmarkCommand RMistral Small 3.2
LMArena Text1187—
LMArena Creative Writing1170—
EQ-Bench Creative Writing—1255
LMArena Multi-Turn1163—
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Mistral Small 3.2?

Command R and Mistral Small 3.2 score almost the same on the Noometry Index (31.4 vs 31.2), so choose on price, context window or the category you care about most.

Which is cheaper, Command R or Mistral Small 3.2?

Mistral Small 3.2 is cheaper. It lists at $0.0938 per million input tokens and $0.25 per million output tokens; Command R lists at $0.15 and $0.60.

Which has the bigger context window?

Mistral Small 3.2 does, with 256K tokens against 128K.

How many benchmarks do Command R and Mistral Small 3.2 share?

0 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Mistral Small 3.2 has 6.

Related comparisons

Go deeper