Model comparison
Granite 4.0 Micro vs Wizardlm 13b
Wizardlm 13b is the stronger model overall, scoring 31.4 to 29.0 on the Noometry Index.
Last verified . 0 shared benchmarks.
Summary
- The widest gap is in math, where Wizardlm 13b leads 30.2 to 12.0.
Side by side
| Granite 4.0 Micro | Wizardlm 13b | |
|---|---|---|
| Provider | IBM | Microsoft |
| Noometry Index | 29.0 | 31.4 |
| Released | 2025-10-02 | — |
| Weights | Open | Open |
| Context window | 131K | — |
| Max output | 118K | — |
| Input $ / M tokens | $0.017 | — |
| Output $ / M tokens | $0.11 | — |
| Results tracked | 8 | 10 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Not comparable
Granite 4.0 Micro: —, Wizardlm 13b: 30.1 (#298)
| Benchmark | Granite 4.0 Micro | Wizardlm 13b |
|---|---|---|
| LMArena Coding | — | 1035 |
Reasoning Too close to call
Granite 4.0 Micro: 19.2 (#265), Wizardlm 13b: 19.4 (#259)
| Benchmark | Granite 4.0 Micro | Wizardlm 13b |
|---|---|---|
| Chess Puzzles | 0% | — |
| LMArena Hard Prompts | — | 1018 |
Math Wizardlm 13b leads
Granite 4.0 Micro: 12.0 (#307), Wizardlm 13b: 30.2 (#238)
| Benchmark | Granite 4.0 Micro | Wizardlm 13b |
|---|---|---|
| OTIS Mock AIME 2024-2025 | 2.8% | — |
| Omni-MATH | 20.9% | — |
| LMArena Math | — | 1017 |
Knowledge Not comparable
Granite 4.0 Micro: 9.9 (#304), Wizardlm 13b: —
| Benchmark | Granite 4.0 Micro | Wizardlm 13b |
|---|---|---|
| GPQA Diamond | 28.3% | — |
| MMLU-Pro | 39.5% | — |
| GPQA (HELM) | 30.7% | — |
Multilingual Not comparable
Granite 4.0 Micro: —, Wizardlm 13b: 27.1 (#277)
| Benchmark | Granite 4.0 Micro | Wizardlm 13b |
|---|---|---|
| LMArena Non-English | — | 1034 |
| LMArena Chinese | — | 1023 |
Instruction Following Granite 4.0 Micro leads
Granite 4.0 Micro: 69.9 (#169), Wizardlm 13b: 53.5 (#285)
| Benchmark | Granite 4.0 Micro | Wizardlm 13b |
|---|---|---|
| IFEval | 84.9% | — |
| LMArena Instruction Following | — | 1048 |
Long Context Not comparable
Granite 4.0 Micro: —, Wizardlm 13b: 32.0 (#273)
| Benchmark | Granite 4.0 Micro | Wizardlm 13b |
|---|---|---|
| LMArena Longer Query | — | 1054 |
Writing & Preference Granite 4.0 Micro leads
Granite 4.0 Micro: 46.7 (#216), Wizardlm 13b: 30.5 (#287)
| Benchmark | Granite 4.0 Micro | Wizardlm 13b |
|---|---|---|
| LMArena Text | — | 1077 |
| LMArena Creative Writing | — | 1091 |
| WildBench | 67% | — |
| LMArena Multi-Turn | — | 1047 |
Frequently asked questions
Is Granite 4.0 Micro better than Wizardlm 13b?
Wizardlm 13b is the stronger model overall, scoring 31.4 to 29.0 on the Noometry Index.
How many benchmarks do Granite 4.0 Micro and Wizardlm 13b share?
0 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Wizardlm 13b has 10.