Model comparison

Mistral Medium 3.5 vs Wizardlm 70b

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 33.0 on the Noometry Index.

Last verified . 12 shared benchmarks.

Mistral Medium 3.5 Mistral AI

40.2

Rank #152 Confirmed

Wizardlm 70b Microsoft

33.0

Rank #249 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Mistral Medium 3.5 scores higher in 6 categories and Wizardlm 70b in 1 category; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mistral Medium 3.5 leads 58.5 to 34.8.

Side by side

Mistral Medium 3.5 and Wizardlm 70b specifications
Mistral Medium 3.5Wizardlm 70b
ProviderMistral AIMicrosoft
Noometry Index40.233.0
Released——
WeightsOpenOpen
Context window262K—
Max output210K—
Input $ / M tokens$1.50—
Output $ / M tokens$7.50—
Results tracked2212

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Medium 3.5 leads

Mistral Medium 3.5: 36.0 (#213), Wizardlm 70b: 31.4 (#285)

Coding benchmarks
BenchmarkMistral Medium 3.5Wizardlm 70b
LMArena Coding14611081
LMArena WebDev1264—

Reasoning Wizardlm 70b leads

Mistral Medium 3.5: 17.3 (#295), Wizardlm 70b: 20.7 (#234)

Reasoning benchmarks
BenchmarkMistral Medium 3.5Wizardlm 70b
LMArena Hard Prompts14361079
Kagi LLM Benchmark41.4%—
NYT Connections (extended)12.9%—
Epoch Capabilities Index141.35—

Math Mistral Medium 3.5 leads

Mistral Medium 3.5: 39.1 (#113), Wizardlm 70b: 32.2 (#218)

Math benchmarks
BenchmarkMistral Medium 3.5Wizardlm 70b
LMArena Math14311116

Knowledge Not comparable

Mistral Medium 3.5: 40.0 (#126), Wizardlm 70b: —

Knowledge benchmarks
BenchmarkMistral Medium 3.5Wizardlm 70b
LMArena Expert1432—

Multimodal Not comparable

Mistral Medium 3.5: 38.3 (#65), Wizardlm 70b: —

Multimodal benchmarks
BenchmarkMistral Medium 3.5Wizardlm 70b
LMArena Vision1223—

Multilingual Mistral Medium 3.5 leads

Mistral Medium 3.5: 51.9 (#100), Wizardlm 70b: 29.6 (#265)

Multilingual benchmarks
BenchmarkMistral Medium 3.5Wizardlm 70b
LMArena Non-English14041078
LMArena Chinese14421052
LMArena German14511083
LMArena Russian13951155
LMArena French1448—
LMArena Korean1385—
LMArena Spanish1409—

Instruction Following Mistral Medium 3.5 leads

Mistral Medium 3.5: 74.6 (#90), Wizardlm 70b: 56.3 (#273)

Instruction Following benchmarks
BenchmarkMistral Medium 3.5Wizardlm 70b
LMArena Instruction Following14151093

Long Context Mistral Medium 3.5 leads

Mistral Medium 3.5: 43.2 (#103), Wizardlm 70b: 33.3 (#263)

Long Context benchmarks
BenchmarkMistral Medium 3.5Wizardlm 70b
LMArena Longer Query14151097

Writing & Preference Mistral Medium 3.5 leads

Mistral Medium 3.5: 58.5 (#117), Wizardlm 70b: 34.8 (#269)

Writing & Preference benchmarks
BenchmarkMistral Medium 3.5Wizardlm 70b
LMArena Text14211120
LMArena Creative Writing13741149
LMArena Multi-Turn14231108
EQ-Bench 4993—

Frequently asked questions

Is Mistral Medium 3.5 better than Wizardlm 70b?

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 33.0 on the Noometry Index.

Is Mistral Medium 3.5 or Wizardlm 70b better for coding?

Mistral Medium 3.5 scores higher on coding benchmarks: 36.0 versus 31.4 in the Noometry coding category.

How many benchmarks do Mistral Medium 3.5 and Wizardlm 70b share?

12 benchmarks have published results for both models. Mistral Medium 3.5 has 22 scored results on Noometry and Wizardlm 70b has 12.

Related comparisons

Go deeper