Model comparison

Mercury vs Wizardlm 13b

Mercury is the stronger model overall, scoring 37.6 to 31.4 on the Noometry Index.

Last verified . 8 shared benchmarks.

Mercury Inception

37.6

Rank #199 Confirmed

Wizardlm 13b Microsoft

31.4

Rank #274 Confirmed

Summary

  • They share 8 benchmarks with published results for both. Mercury scores higher in 5 categories and Wizardlm 13b in 1 category; 6 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mercury leads 46.2 to 30.5.
  • Wizardlm 13b has downloadable open weights; the other is API-only.

Side by side

Mercury and Wizardlm 13b specifications
MercuryWizardlm 13b
ProviderInceptionMicrosoft
Noometry Index37.631.4
Released——
WeightsProprietaryOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked910

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mercury leads

Mercury: 38.7 (#170), Wizardlm 13b: 30.1 (#298)

Coding benchmarks
BenchmarkMercuryWizardlm 13b
LMArena Coding13221035

Reasoning Wizardlm 13b leads

Mercury: 17.5 (#293), Wizardlm 13b: 19.4 (#259)

Reasoning benchmarks
BenchmarkMercuryWizardlm 13b
LMArena Hard Prompts12851018
Kagi LLM Benchmark21.6%—

Math Not comparable

Mercury: —, Wizardlm 13b: 30.2 (#238)

Math benchmarks
BenchmarkMercuryWizardlm 13b
LMArena Math—1017

Multilingual Mercury leads

Mercury: 41.6 (#206), Wizardlm 13b: 27.1 (#277)

Multilingual benchmarks
BenchmarkMercuryWizardlm 13b
LMArena Non-English12601034
LMArena Chinese—1023

Instruction Following Mercury leads

Mercury: 65.2 (#224), Wizardlm 13b: 53.5 (#285)

Instruction Following benchmarks
BenchmarkMercuryWizardlm 13b
LMArena Instruction Following12391048

Long Context Mercury leads

Mercury: 38.4 (#198), Wizardlm 13b: 32.0 (#273)

Long Context benchmarks
BenchmarkMercuryWizardlm 13b
LMArena Longer Query12661054

Writing & Preference Mercury leads

Mercury: 46.2 (#221), Wizardlm 13b: 30.5 (#287)

Writing & Preference benchmarks
BenchmarkMercuryWizardlm 13b
LMArena Text12821077
LMArena Creative Writing11911091
LMArena Multi-Turn12821047

Frequently asked questions

Is Mercury better than Wizardlm 13b?

Mercury is the stronger model overall, scoring 37.6 to 31.4 on the Noometry Index.

Is Mercury or Wizardlm 13b better for coding?

Mercury scores higher on coding benchmarks: 38.7 versus 30.1 in the Noometry coding category.

How many benchmarks do Mercury and Wizardlm 13b share?

8 benchmarks have published results for both models. Mercury has 9 scored results on Noometry and Wizardlm 13b has 10.

Related comparisons

Go deeper