Model comparison

Mistral vs Mistral Medium 3.5

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 29.9 on the Noometry Index.

Last verified . 16 shared benchmarks.

Mistral Mistral AI

29.9

Rank #303 Confirmed

Mistral Medium 3.5 Mistral AI

40.2

Rank #152 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Mistral scores higher in 1 category and Mistral Medium 3.5 in 7 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Mistral Medium 3.5 leads 40.0 to 16.6.
  • Mistral Medium 3.5 has downloadable open weights; the other is API-only.

Side by side

Mistral and Mistral Medium 3.5 specifications
MistralMistral Medium 3.5
ProviderMistral AIMistral AI
Noometry Index29.940.2
Released——
WeightsProprietaryOpen
Context window—262K
Max output—210K
Input $ / M tokens—$1.50
Output $ / M tokens—$7.50
Results tracked2222

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Medium 3.5 leads

Mistral: 33.8 (#250), Mistral Medium 3.5: 36.0 (#213)

Coding benchmarks
BenchmarkMistralMistral Medium 3.5
LMArena Coding11621461
LMArena WebDev—1264

Reasoning Mistral leads

Mistral: 22.2 (#200), Mistral Medium 3.5: 17.3 (#295)

Reasoning benchmarks
BenchmarkMistralMistral Medium 3.5
LMArena Hard Prompts11491436
Kagi LLM Benchmark—41.4%
NYT Connections (extended)—12.9%
Epoch Capabilities Index—141.35

Math Mistral Medium 3.5 leads

Mistral: 22.3 (#278), Mistral Medium 3.5: 39.1 (#113)

Math benchmarks
BenchmarkMistralMistral Medium 3.5
LMArena Math11801431
Omni-MATH7.2%—

Knowledge Mistral Medium 3.5 leads

Mistral: 16.6 (#288), Mistral Medium 3.5: 40.0 (#126)

Knowledge benchmarks
BenchmarkMistralMistral Medium 3.5
LMArena Expert11251432
MMLU-Pro27.7%—
GPQA (HELM)30.3%—

Multimodal Not comparable

Mistral: —, Mistral Medium 3.5: 38.3 (#65)

Multimodal benchmarks
BenchmarkMistralMistral Medium 3.5
LMArena Vision—1223

Multilingual Mistral Medium 3.5 leads

Mistral: 32.8 (#254), Mistral Medium 3.5: 51.9 (#100)

Multilingual benchmarks
BenchmarkMistralMistral Medium 3.5
LMArena Non-English11291404
LMArena Chinese11091442
LMArena French11801448
LMArena German11551451
LMArena Korean10321385
LMArena Russian11681395
LMArena Spanish11431409
LMArena Japanese1013—

Instruction Following Mistral Medium 3.5 leads

Mistral: 52.6 (#288), Mistral Medium 3.5: 74.6 (#90)

Instruction Following benchmarks
BenchmarkMistralMistral Medium 3.5
LMArena Instruction Following11521415
IFEval56.8%—

Long Context Mistral Medium 3.5 leads

Mistral: 35.0 (#245), Mistral Medium 3.5: 43.2 (#103)

Long Context benchmarks
BenchmarkMistralMistral Medium 3.5
LMArena Longer Query11531415

Writing & Preference Mistral Medium 3.5 leads

Mistral: 37.0 (#260), Mistral Medium 3.5: 58.5 (#117)

Writing & Preference benchmarks
BenchmarkMistralMistral Medium 3.5
LMArena Text11651421
LMArena Creative Writing11581374
LMArena Multi-Turn11471423
WildBench66%—
EQ-Bench 4—993

Frequently asked questions

Is Mistral better than Mistral Medium 3.5?

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 29.9 on the Noometry Index.

Is Mistral or Mistral Medium 3.5 better for coding?

Mistral Medium 3.5 scores higher on coding benchmarks: 36.0 versus 33.8 in the Noometry coding category.

How many benchmarks do Mistral and Mistral Medium 3.5 share?

16 benchmarks have published results for both models. Mistral has 22 scored results on Noometry and Mistral Medium 3.5 has 22.

Related comparisons

Go deeper