Model comparison

Llama 3.3 Nemotron 49b Super v1 vs Wizardlm 13b

Llama 3.3 Nemotron 49b Super v1 is the stronger model overall, scoring 40.1 to 31.4 on the Noometry Index.

Last verified . 9 shared benchmarks.

Wizardlm 13b Microsoft

31.4

Rank #274 Confirmed

Summary

  • They share 9 benchmarks with published results for both. Llama 3.3 Nemotron 49b Super v1 scores higher in 6 categories and Wizardlm 13b in 0 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Llama 3.3 Nemotron 49b Super v1 leads 50.8 to 30.5.

Side by side

Llama 3.3 Nemotron 49b Super v1 and Wizardlm 13b specifications
Llama 3.3 Nemotron 49b Super v1Wizardlm 13b
ProviderNVIDIAMicrosoft
Noometry Index40.131.4
Released——
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1010

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Llama 3.3 Nemotron 49b Super v1 leads

Llama 3.3 Nemotron 49b Super v1: 37.9 (#186), Wizardlm 13b: 30.1 (#298)

Coding benchmarks
BenchmarkLlama 3.3 Nemotron 49b Super v1Wizardlm 13b
LMArena Coding12961035

Reasoning Llama 3.3 Nemotron 49b Super v1 leads

Llama 3.3 Nemotron 49b Super v1: 26.2 (#135), Wizardlm 13b: 19.4 (#259)

Reasoning benchmarks
BenchmarkLlama 3.3 Nemotron 49b Super v1Wizardlm 13b
LMArena Hard Prompts13111018

Math Not comparable

Llama 3.3 Nemotron 49b Super v1: —, Wizardlm 13b: 30.2 (#238)

Math benchmarks
BenchmarkLlama 3.3 Nemotron 49b Super v1Wizardlm 13b
LMArena Math—1017

Multilingual Llama 3.3 Nemotron 49b Super v1 leads

Llama 3.3 Nemotron 49b Super v1: 41.1 (#211), Wizardlm 13b: 27.1 (#277)

Multilingual benchmarks
BenchmarkLlama 3.3 Nemotron 49b Super v1Wizardlm 13b
LMArena Non-English12531034
LMArena Chinese12771023
LMArena Russian1269—

Instruction Following Llama 3.3 Nemotron 49b Super v1 leads

Llama 3.3 Nemotron 49b Super v1: 68.3 (#189), Wizardlm 13b: 53.5 (#285)

Instruction Following benchmarks
BenchmarkLlama 3.3 Nemotron 49b Super v1Wizardlm 13b
LMArena Instruction Following12931048

Long Context Llama 3.3 Nemotron 49b Super v1 leads

Llama 3.3 Nemotron 49b Super v1: 39.5 (#176), Wizardlm 13b: 32.0 (#273)

Long Context benchmarks
BenchmarkLlama 3.3 Nemotron 49b Super v1Wizardlm 13b
LMArena Longer Query12991054

Writing & Preference Llama 3.3 Nemotron 49b Super v1 leads

Llama 3.3 Nemotron 49b Super v1: 50.8 (#179), Wizardlm 13b: 30.5 (#287)

Writing & Preference benchmarks
BenchmarkLlama 3.3 Nemotron 49b Super v1Wizardlm 13b
LMArena Text13081077
LMArena Creative Writing12881091
LMArena Multi-Turn13151047

Frequently asked questions

Is Llama 3.3 Nemotron 49b Super v1 better than Wizardlm 13b?

Llama 3.3 Nemotron 49b Super v1 is the stronger model overall, scoring 40.1 to 31.4 on the Noometry Index.

Is Llama 3.3 Nemotron 49b Super v1 or Wizardlm 13b better for coding?

Llama 3.3 Nemotron 49b Super v1 scores higher on coding benchmarks: 37.9 versus 30.1 in the Noometry coding category.

How many benchmarks do Llama 3.3 Nemotron 49b Super v1 and Wizardlm 13b share?

9 benchmarks have published results for both models. Llama 3.3 Nemotron 49b Super v1 has 10 scored results on Noometry and Wizardlm 13b has 10.

Related comparisons

Go deeper