Model comparison

Command R vs Mistral Large 3

Mistral Large 3 is the stronger model overall, scoring 39.1 to 31.4 on the Noometry Index.

Last verified . 17 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Mistral Large 3 Mistral AI

39.1

Rank #176 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Command R scores higher in 0 categories and Mistral Large 3 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mistral Large 3 leads 60.0 to 38.2.
  • Command R is cheaper at $0.15 / $0.60 per million input/output tokens, against $0.25 / $0.75 for Mistral Large 3.
  • Mistral Large 3 accepts more context: 262K tokens versus 128K.

Side by side

Command R and Mistral Large 3 specifications
Command RMistral Large 3
ProviderCohereMistral AI
Noometry Index31.439.1
Released2024-08-302025-12-02
WeightsOpenOpen
Context window128K262K
Max output4K8K
Input $ / M tokens$0.15$0.25
Output $ / M tokens$0.60$0.75
Results tracked2924

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Large 3 leads

Command R: 29.3 (#306), Mistral Large 3: 34.4 (#237)

Coding benchmarks
BenchmarkCommand RMistral Large 3
LMArena Coding11691448
LMArena WebDev—1230
BigCodeBench Instruct37.1%—
LiveBench Coding17.9%—
BigCodeBench Complete45.2%—

Reasoning Mistral Large 3 leads

Command R: 13.8 (#331), Mistral Large 3: 15.2 (#319)

Reasoning benchmarks
BenchmarkCommand RMistral Large 3
LMArena Hard Prompts11641429
Kagi LLM Benchmark—50.9%
NYT Connections (extended)—7.5%
Thematic Generalization—23%
LiveBench Reasoning21.9%—
DTBench46.4%—
LiveBench Data Analysis33.3%—
LMCA9.2%—
LiveBench27.5%—

Math Mistral Large 3 leads

Command R: 28.0 (#246), Mistral Large 3: 38.7 (#129)

Math benchmarks
BenchmarkCommand RMistral Large 3
LMArena Math11551414
LiveBench Math19.4%—

Knowledge Mistral Large 3 leads

Command R: 31.0 (#221), Mistral Large 3: 36.0 (#177)

Knowledge benchmarks
BenchmarkCommand RMistral Large 3
LMArena Expert11381421
Vectara Hallucination Rate—14.5%
MMLU65.2%—

Multimodal Not comparable

Command R: —, Mistral Large 3: 38.2 (#66)

Multimodal benchmarks
BenchmarkCommand RMistral Large 3
LMArena Vision—1221

Multilingual Mistral Large 3 leads

Command R: 35.7 (#245), Mistral Large 3: 52.5 (#84)

Multilingual benchmarks
BenchmarkCommand RMistral Large 3
LMArena Non-English11741413
LMArena Chinese11821447
LMArena French11621455
LMArena German11761437
LMArena Japanese11431394
LMArena Korean11631384
LMArena Russian11741411
LMArena Spanish11511440

Instruction Following Mistral Large 3 leads

Command R: 58.1 (#261), Mistral Large 3: 74.0 (#108)

Instruction Following benchmarks
BenchmarkCommand RMistral Large 3
LMArena Instruction Following11671403
LiveBench Instruction Following55.6%—

Long Context Mistral Large 3 leads

Command R: 36.3 (#231), Mistral Large 3: 43.1 (#105)

Long Context benchmarks
BenchmarkCommand RMistral Large 3
LMArena Longer Query11981413

Writing & Preference Mistral Large 3 leads

Command R: 38.2 (#254), Mistral Large 3: 60.0 (#101)

Writing & Preference benchmarks
BenchmarkCommand RMistral Large 3
LMArena Text11871428
LMArena Creative Writing11701386
LMArena Multi-Turn11631429
EQ-Bench Creative Writing—1412
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Mistral Large 3?

Mistral Large 3 is the stronger model overall, scoring 39.1 to 31.4 on the Noometry Index.

Which is cheaper, Command R or Mistral Large 3?

Command R is cheaper. It lists at $0.15 per million input tokens and $0.60 per million output tokens; Mistral Large 3 lists at $0.25 and $0.75.

Is Command R or Mistral Large 3 better for coding?

Mistral Large 3 scores higher on coding benchmarks: 34.4 versus 29.3 in the Noometry coding category.

Which has the bigger context window?

Mistral Large 3 does, with 262K tokens against 128K.

How many benchmarks do Command R and Mistral Large 3 share?

17 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Mistral Large 3 has 24.

Related comparisons

Go deeper