Model comparison

Magistral Small vs Wizardlm 13b

Wizardlm 13b is the stronger model overall, scoring 31.4 to 30.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Magistral Small Mistral AI

30.2

Rank #296 Confirmed

Wizardlm 13b Microsoft

31.4

Rank #274 Confirmed

Summary

  • The widest gap is in reasoning, where Wizardlm 13b leads 19.4 to 6.8.

Side by side

Magistral Small and Wizardlm 13b specifications
Magistral SmallWizardlm 13b
ProviderMistral AIMicrosoft
Noometry Index30.231.4
Released2025-06-10—
WeightsOpenOpen
Context window128K—
Max output40K—
Input $ / M tokens$0.50—
Output $ / M tokens$1.50—
Results tracked1010

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Magistral Small leads

Magistral Small: 38.4 (#176), Wizardlm 13b: 30.1 (#298)

Coding benchmarks
BenchmarkMagistral SmallWizardlm 13b
SciCode35.2%—
LMArena Coding—1035

Reasoning Wizardlm 13b leads

Magistral Small: 6.8 (#350), Wizardlm 13b: 19.4 (#259)

Reasoning benchmarks
BenchmarkMagistral SmallWizardlm 13b
ARC-AGI-20%—
Kagi LLM Benchmark6.3%—
ARC-AGI-15%—
CritPt0.3%—
Chess Puzzles3%—
LMArena Hard Prompts—1018
DTBench61.3%—
Epoch Capabilities Index133.19—

Math Wizardlm 13b leads

Magistral Small: 26.2 (#261), Wizardlm 13b: 30.2 (#238)

Math benchmarks
BenchmarkMagistral SmallWizardlm 13b
OTIS Mock AIME 2024-202530%—
LMArena Math—1017

Knowledge Not comparable

Magistral Small: 30.9 (#223), Wizardlm 13b: —

Knowledge benchmarks
BenchmarkMagistral SmallWizardlm 13b
GPQA Diamond56.1%—

Multilingual Not comparable

Magistral Small: —, Wizardlm 13b: 27.1 (#277)

Multilingual benchmarks
BenchmarkMagistral SmallWizardlm 13b
LMArena Non-English—1034
LMArena Chinese—1023

Instruction Following Not comparable

Magistral Small: —, Wizardlm 13b: 53.5 (#285)

Instruction Following benchmarks
BenchmarkMagistral SmallWizardlm 13b
LMArena Instruction Following—1048

Long Context Not comparable

Magistral Small: —, Wizardlm 13b: 32.0 (#273)

Long Context benchmarks
BenchmarkMagistral SmallWizardlm 13b
LMArena Longer Query—1054

Writing & Preference Not comparable

Magistral Small: —, Wizardlm 13b: 30.5 (#287)

Writing & Preference benchmarks
BenchmarkMagistral SmallWizardlm 13b
LMArena Text—1077
LMArena Creative Writing—1091
LMArena Multi-Turn—1047

Frequently asked questions

Is Magistral Small better than Wizardlm 13b?

Wizardlm 13b is the stronger model overall, scoring 31.4 to 30.2 on the Noometry Index.

Is Magistral Small or Wizardlm 13b better for coding?

Magistral Small scores higher on coding benchmarks: 38.4 versus 30.1 in the Noometry coding category.

How many benchmarks do Magistral Small and Wizardlm 13b share?

0 benchmarks have published results for both models. Magistral Small has 10 scored results on Noometry and Wizardlm 13b has 10.

Related comparisons

Go deeper