Model comparison

Granite 4.2 30b vs Wizardlm 13b

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 31.4 on the Noometry Index.

Last verified . 9 shared benchmarks.

Granite 4.2 30b IBM

41.8

Rank #130 Confirmed

Wizardlm 13b Microsoft

31.4

Rank #274 Confirmed

Summary

  • They share 9 benchmarks with published results for both. Granite 4.2 30b scores higher in 6 categories and Wizardlm 13b in 0 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Granite 4.2 30b leads 53.8 to 30.5.

Side by side

Granite 4.2 30b and Wizardlm 13b specifications
Granite 4.2 30bWizardlm 13b
ProviderIBMMicrosoft
Noometry Index41.831.4
Released——
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1110

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Granite 4.2 30b leads

Granite 4.2 30b: 41.0 (#126), Wizardlm 13b: 30.1 (#298)

Coding benchmarks
BenchmarkGranite 4.2 30bWizardlm 13b
LMArena Coding13961035

Reasoning Granite 4.2 30b leads

Granite 4.2 30b: 27.8 (#112), Wizardlm 13b: 19.4 (#259)

Reasoning benchmarks
BenchmarkGranite 4.2 30bWizardlm 13b
LMArena Hard Prompts13741018

Math Not comparable

Granite 4.2 30b: —, Wizardlm 13b: 30.2 (#238)

Math benchmarks
BenchmarkGranite 4.2 30bWizardlm 13b
LMArena Math—1017

Knowledge Not comparable

Granite 4.2 30b: 39.1 (#138), Wizardlm 13b: —

Knowledge benchmarks
BenchmarkGranite 4.2 30bWizardlm 13b
LMArena Expert1406—

Multilingual Granite 4.2 30b leads

Granite 4.2 30b: 47.3 (#151), Wizardlm 13b: 27.1 (#277)

Multilingual benchmarks
BenchmarkGranite 4.2 30bWizardlm 13b
LMArena Non-English13401034
LMArena Chinese14141023
LMArena Russian1343—

Instruction Following Granite 4.2 30b leads

Granite 4.2 30b: 71.2 (#155), Wizardlm 13b: 53.5 (#285)

Instruction Following benchmarks
BenchmarkGranite 4.2 30bWizardlm 13b
LMArena Instruction Following13471048

Long Context Granite 4.2 30b leads

Granite 4.2 30b: 41.4 (#140), Wizardlm 13b: 32.0 (#273)

Long Context benchmarks
BenchmarkGranite 4.2 30bWizardlm 13b
LMArena Longer Query13591054

Writing & Preference Granite 4.2 30b leads

Granite 4.2 30b: 53.8 (#156), Wizardlm 13b: 30.5 (#287)

Writing & Preference benchmarks
BenchmarkGranite 4.2 30bWizardlm 13b
LMArena Text13611077
LMArena Creative Writing12881091
LMArena Multi-Turn13391047

Frequently asked questions

Is Granite 4.2 30b better than Wizardlm 13b?

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 31.4 on the Noometry Index.

Is Granite 4.2 30b or Wizardlm 13b better for coding?

Granite 4.2 30b scores higher on coding benchmarks: 41.0 versus 30.1 in the Noometry coding category.

How many benchmarks do Granite 4.2 30b and Wizardlm 13b share?

9 benchmarks have published results for both models. Granite 4.2 30b has 11 scored results on Noometry and Wizardlm 13b has 10.

Related comparisons

Go deeper