Model comparison

Magistral Medium vs Wizardlm 13b

Magistral Medium is the stronger model overall, scoring 35.2 to 31.4 on the Noometry Index.

Last verified . 10 shared benchmarks.

Magistral Medium Mistral AI

35.2

Rank #227 Confirmed

Wizardlm 13b Microsoft

31.4

Rank #274 Confirmed

Summary

  • They share 10 benchmarks with published results for both. Magistral Medium scores higher in 6 categories and Wizardlm 13b in 1 category; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Magistral Medium leads 46.3 to 30.5.

Side by side

Magistral Medium and Wizardlm 13b specifications
Magistral MediumWizardlm 13b
ProviderMistral AIMicrosoft
Noometry Index35.231.4
Released2025-03-17—
WeightsOpenOpen
Context window262K—
Max output16K—
Input $ / M tokens$2—
Output $ / M tokens$5—
Results tracked2210

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Magistral Medium leads

Magistral Medium: 39.1 (#161), Wizardlm 13b: 30.1 (#298)

Coding benchmarks
BenchmarkMagistral MediumWizardlm 13b
LMArena Coding13191035
SciCode39.2%—

Reasoning Wizardlm 13b leads

Magistral Medium: 8.6 (#348), Wizardlm 13b: 19.4 (#259)

Reasoning benchmarks
BenchmarkMagistral MediumWizardlm 13b
LMArena Hard Prompts12671018
ARC-AGI-20%—
Kagi LLM Benchmark16.2%—
ARC-AGI-16.1%—
CritPt0.3%—

Math Magistral Medium leads

Magistral Medium: 35.1 (#189), Wizardlm 13b: 30.2 (#238)

Math benchmarks
BenchmarkMagistral MediumWizardlm 13b
LMArena Math12501017

Knowledge Not comparable

Magistral Medium: 33.5 (#202), Wizardlm 13b: —

Knowledge benchmarks
BenchmarkMagistral MediumWizardlm 13b
LMArena Expert1223—

Multilingual Magistral Medium leads

Magistral Medium: 39.6 (#224), Wizardlm 13b: 27.1 (#277)

Multilingual benchmarks
BenchmarkMagistral MediumWizardlm 13b
LMArena Non-English12321034
LMArena Chinese12271023
LMArena French1267—
LMArena German1248—
LMArena Japanese1175—
LMArena Korean1125—
LMArena Russian1224—
LMArena Spanish1271—

Instruction Following Magistral Medium leads

Magistral Medium: 66.0 (#211), Wizardlm 13b: 53.5 (#285)

Instruction Following benchmarks
BenchmarkMagistral MediumWizardlm 13b
LMArena Instruction Following12541048

Long Context Magistral Medium leads

Magistral Medium: 39.3 (#183), Wizardlm 13b: 32.0 (#273)

Long Context benchmarks
BenchmarkMagistral MediumWizardlm 13b
LMArena Longer Query12951054

Writing & Preference Magistral Medium leads

Magistral Medium: 46.3 (#219), Wizardlm 13b: 30.5 (#287)

Writing & Preference benchmarks
BenchmarkMagistral MediumWizardlm 13b
LMArena Text12551077
LMArena Creative Writing12451091
LMArena Multi-Turn12751047

Frequently asked questions

Is Magistral Medium better than Wizardlm 13b?

Magistral Medium is the stronger model overall, scoring 35.2 to 31.4 on the Noometry Index.

Is Magistral Medium or Wizardlm 13b better for coding?

Magistral Medium scores higher on coding benchmarks: 39.1 versus 30.1 in the Noometry coding category.

How many benchmarks do Magistral Medium and Wizardlm 13b share?

10 benchmarks have published results for both models. Magistral Medium has 22 scored results on Noometry and Wizardlm 13b has 10.

Related comparisons

Go deeper