Model comparison

Magistral Medium vs Wizardlm 70b

Magistral Medium is the stronger model overall, scoring 35.2 to 33.0 on the Noometry Index.

Last verified . 12 shared benchmarks.

Magistral Medium Mistral AI

35.2

Rank #227 Confirmed

Wizardlm 70b Microsoft

33.0

Rank #249 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Magistral Medium scores higher in 6 categories and Wizardlm 70b in 1 category; 7 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Wizardlm 70b leads 20.7 to 8.6.

Side by side

Magistral Medium and Wizardlm 70b specifications
Magistral MediumWizardlm 70b
ProviderMistral AIMicrosoft
Noometry Index35.233.0
Released2025-03-17—
WeightsOpenOpen
Context window262K—
Max output16K—
Input $ / M tokens$2—
Output $ / M tokens$5—
Results tracked2212

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Magistral Medium leads

Magistral Medium: 39.1 (#161), Wizardlm 70b: 31.4 (#285)

Coding benchmarks
BenchmarkMagistral MediumWizardlm 70b
LMArena Coding13191081
SciCode39.2%—

Reasoning Wizardlm 70b leads

Magistral Medium: 8.6 (#348), Wizardlm 70b: 20.7 (#234)

Reasoning benchmarks
BenchmarkMagistral MediumWizardlm 70b
LMArena Hard Prompts12671079
ARC-AGI-20%—
Kagi LLM Benchmark16.2%—
ARC-AGI-16.1%—
CritPt0.3%—

Math Magistral Medium leads

Magistral Medium: 35.1 (#189), Wizardlm 70b: 32.2 (#218)

Math benchmarks
BenchmarkMagistral MediumWizardlm 70b
LMArena Math12501116

Knowledge Not comparable

Magistral Medium: 33.5 (#202), Wizardlm 70b: —

Knowledge benchmarks
BenchmarkMagistral MediumWizardlm 70b
LMArena Expert1223—

Multilingual Magistral Medium leads

Magistral Medium: 39.6 (#224), Wizardlm 70b: 29.6 (#265)

Multilingual benchmarks
BenchmarkMagistral MediumWizardlm 70b
LMArena Non-English12321078
LMArena Chinese12271052
LMArena German12481083
LMArena Russian12241155
LMArena French1267—
LMArena Japanese1175—
LMArena Korean1125—
LMArena Spanish1271—

Instruction Following Magistral Medium leads

Magistral Medium: 66.0 (#211), Wizardlm 70b: 56.3 (#273)

Instruction Following benchmarks
BenchmarkMagistral MediumWizardlm 70b
LMArena Instruction Following12541093

Long Context Magistral Medium leads

Magistral Medium: 39.3 (#183), Wizardlm 70b: 33.3 (#263)

Long Context benchmarks
BenchmarkMagistral MediumWizardlm 70b
LMArena Longer Query12951097

Writing & Preference Magistral Medium leads

Magistral Medium: 46.3 (#219), Wizardlm 70b: 34.8 (#269)

Writing & Preference benchmarks
BenchmarkMagistral MediumWizardlm 70b
LMArena Text12551120
LMArena Creative Writing12451149
LMArena Multi-Turn12751108

Frequently asked questions

Is Magistral Medium better than Wizardlm 70b?

Magistral Medium is the stronger model overall, scoring 35.2 to 33.0 on the Noometry Index.

Is Magistral Medium or Wizardlm 70b better for coding?

Magistral Medium scores higher on coding benchmarks: 39.1 versus 31.4 in the Noometry coding category.

How many benchmarks do Magistral Medium and Wizardlm 70b share?

12 benchmarks have published results for both models. Magistral Medium has 22 scored results on Noometry and Wizardlm 70b has 12.

Related comparisons

Go deeper