Model comparison

Nvidia Llama 3.3 Nemotron Super 49b v1.5 vs Qwen1.5-32B

Nvidia Llama 3.3 Nemotron Super 49b v1.5 is the stronger model overall, scoring 40.3 to 30.5 on the Noometry Index.

Last verified . 12 shared benchmarks.

Qwen1.5-32B Alibaba (Qwen)

30.5

Rank #293 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Nvidia Llama 3.3 Nemotron Super 49b v1.5 scores higher in 8 categories and Qwen1.5-32B in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads 36.7 to 13.5.

Side by side

Nvidia Llama 3.3 Nemotron Super 49b v1.5 and Qwen1.5-32B specifications
Nvidia Llama 3.3 Nemotron Super 49b v1.5Qwen1.5-32B
ProviderNVIDIAAlibaba (Qwen)
Noometry Index40.330.5
Released2025-07-252024-02-04
WeightsOpenOpen
Context window131K—
Max output131K—
Input $ / M tokens$0.40—
Output $ / M tokens$0.40—
Results tracked1221

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 39.8 (#154), Qwen1.5-32B: 31.7 (#282)

Coding benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen1.5-32B
LMArena Coding13551155
BigCodeBench Instruct—32.3%
BigCodeBench Complete—42%

Reasoning Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 26.8 (#128), Qwen1.5-32B: 21.8 (#212)

Reasoning benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen1.5-32B
LMArena Hard Prompts13361130

Math Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 38.2 (#141), Qwen1.5-32B: 33.0 (#207)

Math benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen1.5-32B
LMArena Math13921155

Knowledge Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 36.7 (#165), Qwen1.5-32B: 13.5 (#296)

Knowledge benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen1.5-32B
LMArena Expert13301126
GPQA Diamond—30.7%
MMLU—74.4%

Multilingual Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 45.5 (#168), Qwen1.5-32B: 31.4 (#259)

Multilingual benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen1.5-32B
LMArena Non-English13161106
LMArena Japanese13001027
LMArena Russian13321073
LMArena Chinese—1177
LMArena French—1101
LMArena German—1058
LMArena Korean—1008
LMArena Spanish—1089

Instruction Following Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 68.6 (#188), Qwen1.5-32B: 57.7 (#265)

Instruction Following benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen1.5-32B
LMArena Instruction Following12991116

Long Context Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 40.0 (#164), Qwen1.5-32B: 34.7 (#246)

Long Context benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen1.5-32B
LMArena Longer Query13151146

Writing & Preference Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 53.1 (#159), Qwen1.5-32B: 34.2 (#271)

Writing & Preference benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen1.5-32B
LMArena Text13381137
LMArena Creative Writing13071083
LMArena Multi-Turn13341140

Frequently asked questions

Is Nvidia Llama 3.3 Nemotron Super 49b v1.5 better than Qwen1.5-32B?

Nvidia Llama 3.3 Nemotron Super 49b v1.5 is the stronger model overall, scoring 40.3 to 30.5 on the Noometry Index.

Is Nvidia Llama 3.3 Nemotron Super 49b v1.5 or Qwen1.5-32B better for coding?

Nvidia Llama 3.3 Nemotron Super 49b v1.5 scores higher on coding benchmarks: 39.8 versus 31.7 in the Noometry coding category.

How many benchmarks do Nvidia Llama 3.3 Nemotron Super 49b v1.5 and Qwen1.5-32B share?

12 benchmarks have published results for both models. Nvidia Llama 3.3 Nemotron Super 49b v1.5 has 12 scored results on Noometry and Qwen1.5-32B has 21.

Related comparisons

Go deeper