Model comparison

Magistral Small vs Wizardlm 70b

Wizardlm 70b is the stronger model overall, scoring 33.0 to 30.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Magistral Small Mistral AI

30.2

Rank #296 Confirmed

Wizardlm 70b Microsoft

33.0

Rank #249 Confirmed

Summary

  • The widest gap is in reasoning, where Wizardlm 70b leads 20.7 to 6.8.

Side by side

Magistral Small and Wizardlm 70b specifications
Magistral SmallWizardlm 70b
ProviderMistral AIMicrosoft
Noometry Index30.233.0
Released2025-06-10—
WeightsOpenOpen
Context window128K—
Max output40K—
Input $ / M tokens$0.50—
Output $ / M tokens$1.50—
Results tracked1012

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Magistral Small leads

Magistral Small: 38.4 (#176), Wizardlm 70b: 31.4 (#285)

Coding benchmarks
BenchmarkMagistral SmallWizardlm 70b
SciCode35.2%—
LMArena Coding—1081

Reasoning Wizardlm 70b leads

Magistral Small: 6.8 (#350), Wizardlm 70b: 20.7 (#234)

Reasoning benchmarks
BenchmarkMagistral SmallWizardlm 70b
ARC-AGI-20%—
Kagi LLM Benchmark6.3%—
ARC-AGI-15%—
CritPt0.3%—
Chess Puzzles3%—
LMArena Hard Prompts—1079
DTBench61.3%—
Epoch Capabilities Index133.19—

Math Wizardlm 70b leads

Magistral Small: 26.2 (#261), Wizardlm 70b: 32.2 (#218)

Math benchmarks
BenchmarkMagistral SmallWizardlm 70b
OTIS Mock AIME 2024-202530%—
LMArena Math—1116

Knowledge Not comparable

Magistral Small: 30.9 (#223), Wizardlm 70b: —

Knowledge benchmarks
BenchmarkMagistral SmallWizardlm 70b
GPQA Diamond56.1%—

Multilingual Not comparable

Magistral Small: —, Wizardlm 70b: 29.6 (#265)

Multilingual benchmarks
BenchmarkMagistral SmallWizardlm 70b
LMArena Non-English—1078
LMArena Chinese—1052
LMArena German—1083
LMArena Russian—1155

Instruction Following Not comparable

Magistral Small: —, Wizardlm 70b: 56.3 (#273)

Instruction Following benchmarks
BenchmarkMagistral SmallWizardlm 70b
LMArena Instruction Following—1093

Long Context Not comparable

Magistral Small: —, Wizardlm 70b: 33.3 (#263)

Long Context benchmarks
BenchmarkMagistral SmallWizardlm 70b
LMArena Longer Query—1097

Writing & Preference Not comparable

Magistral Small: —, Wizardlm 70b: 34.8 (#269)

Writing & Preference benchmarks
BenchmarkMagistral SmallWizardlm 70b
LMArena Text—1120
LMArena Creative Writing—1149
LMArena Multi-Turn—1108

Frequently asked questions

Is Magistral Small better than Wizardlm 70b?

Wizardlm 70b is the stronger model overall, scoring 33.0 to 30.2 on the Noometry Index.

Is Magistral Small or Wizardlm 70b better for coding?

Magistral Small scores higher on coding benchmarks: 38.4 versus 31.4 in the Noometry coding category.

How many benchmarks do Magistral Small and Wizardlm 70b share?

0 benchmarks have published results for both models. Magistral Small has 10 scored results on Noometry and Wizardlm 70b has 12.

Related comparisons

Go deeper