Model comparison

Codestral vs Wizardlm 70b

Wizardlm 70b is the stronger model overall, scoring 33.0 to 30.6 on the Noometry Index.

Last verified . 0 shared benchmarks.

Codestral Mistral AI

30.6

Rank #290 Reported

Wizardlm 70b Microsoft

33.0

Rank #249 Confirmed

Summary

  • The widest gap is in coding, where Wizardlm 70b leads 31.4 to 27.3.
  • Wizardlm 70b has downloadable open weights; the other is API-only.

Side by side

Codestral and Wizardlm 70b specifications
CodestralWizardlm 70b
ProviderMistral AIMicrosoft
Noometry Index30.633.0
Released2024-05-29—
WeightsProprietaryOpen
Context window256K—
Max output8K—
Input $ / M tokens$0.30—
Output $ / M tokens$0.90—
Results tracked712

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Wizardlm 70b leads

Codestral: 27.3 (#321), Wizardlm 70b: 31.4 (#285)

Coding benchmarks
BenchmarkCodestralWizardlm 70b
Aider Polyglot11.1%—
BigCodeBench Instruct41.8%—
LMArena Coding—1081
BigCodeBench Complete52.5%—
ALE-Bench137.78—
HumanEval+73.8%—
MBPP+61.9%—

Reasoning Too close to call

Codestral: 19.8 (#251), Wizardlm 70b: 20.7 (#234)

Reasoning benchmarks
BenchmarkCodestralWizardlm 70b
Kagi LLM Benchmark32.5%—
LMArena Hard Prompts—1079

Math Not comparable

Codestral: —, Wizardlm 70b: 32.2 (#218)

Math benchmarks
BenchmarkCodestralWizardlm 70b
LMArena Math—1116

Multilingual Not comparable

Codestral: —, Wizardlm 70b: 29.6 (#265)

Multilingual benchmarks
BenchmarkCodestralWizardlm 70b
LMArena Non-English—1078
LMArena Chinese—1052
LMArena German—1083
LMArena Russian—1155

Instruction Following Not comparable

Codestral: —, Wizardlm 70b: 56.3 (#273)

Instruction Following benchmarks
BenchmarkCodestralWizardlm 70b
LMArena Instruction Following—1093

Long Context Not comparable

Codestral: —, Wizardlm 70b: 33.3 (#263)

Long Context benchmarks
BenchmarkCodestralWizardlm 70b
LMArena Longer Query—1097

Writing & Preference Not comparable

Codestral: —, Wizardlm 70b: 34.8 (#269)

Writing & Preference benchmarks
BenchmarkCodestralWizardlm 70b
LMArena Text—1120
LMArena Creative Writing—1149
LMArena Multi-Turn—1108

Frequently asked questions

Is Codestral better than Wizardlm 70b?

Wizardlm 70b is the stronger model overall, scoring 33.0 to 30.6 on the Noometry Index.

Is Codestral or Wizardlm 70b better for coding?

Wizardlm 70b scores higher on coding benchmarks: 31.4 versus 27.3 in the Noometry coding category.

How many benchmarks do Codestral and Wizardlm 70b share?

0 benchmarks have published results for both models. Codestral has 7 scored results on Noometry and Wizardlm 70b has 12.

Related comparisons

Go deeper