Model comparison

Nvidia Llama 3.3 Nemotron Super 49b v1.5 vs Qwen3.5 35B-A3B

Qwen3.5 35B-A3B is the stronger model overall, scoring 42.0 to 40.3 on the Noometry Index. Nvidia Llama 3.3 Nemotron Super 49b v1.5 costs 1.7× less per token, which makes it the better buy when Qwen3.5 35B-A3B's lead doesn't matter for your workload.

Last verified . 12 shared benchmarks.

Summary

  • They share 12 benchmarks with published results for both. Nvidia Llama 3.3 Nemotron Super 49b v1.5 scores higher in 2 categories and Qwen3.5 35B-A3B in 6 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Qwen3.5 35B-A3B leads 47.8 to 36.7.
  • Nvidia Llama 3.3 Nemotron Super 49b v1.5 is cheaper at $0.40 / $0.40 per million input/output tokens, against $0.25 / $2 for Qwen3.5 35B-A3B.
  • Qwen3.5 35B-A3B accepts more context: 262K tokens versus 131K.

Side by side

Nvidia Llama 3.3 Nemotron Super 49b v1.5 and Qwen3.5 35B-A3B specifications
Nvidia Llama 3.3 Nemotron Super 49b v1.5Qwen3.5 35B-A3B
ProviderNVIDIAAlibaba (Qwen)
Noometry Index40.342.0
Released2025-07-252026-02-01
WeightsOpenOpen
Context window131K262K
Max output131K66K
Input $ / M tokens$0.40$0.25
Output $ / M tokens$0.40$2
Results tracked1228

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 39.8 (#154), Qwen3.5 35B-A3B: 33.8 (#251)

Coding benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen3.5 35B-A3B
LMArena Coding13551410
LMArena WebDev—1254
SciCode—29.3%

Reasoning Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 26.8 (#128), Qwen3.5 35B-A3B: 24.6 (#161)

Reasoning benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen3.5 35B-A3B
LMArena Hard Prompts13361400
CritPt—0.6%
Chess Puzzles—10%
DTBench—80%
LMCA—29.5%
Epoch Capabilities Index—142.52

Math Qwen3.5 35B-A3B leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 38.2 (#141), Qwen3.5 35B-A3B: 39.9 (#97)

Math benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen3.5 35B-A3B
LMArena Math13921404
MathArena Final-Answer Competitions—56%
OTIS Mock AIME 2024-2025—70%

Knowledge Qwen3.5 35B-A3B leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 36.7 (#165), Qwen3.5 35B-A3B: 47.8 (#79)

Knowledge benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen3.5 35B-A3B
LMArena Expert13301408
GPQA Diamond—83.5%
Vectara Hallucination Rate—10.5%

Multilingual Qwen3.5 35B-A3B leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 45.5 (#168), Qwen3.5 35B-A3B: 50.0 (#127)

Multilingual benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen3.5 35B-A3B
LMArena Non-English13161378
LMArena Japanese13001325
LMArena Russian13321376
LMArena Chinese—1457
LMArena French—1412
LMArena German—1367
LMArena Korean—1356
LMArena Spanish—1392

Instruction Following Qwen3.5 35B-A3B leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 68.6 (#188), Qwen3.5 35B-A3B: 72.8 (#128)

Instruction Following benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen3.5 35B-A3B
LMArena Instruction Following12991379

Long Context Qwen3.5 35B-A3B leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 40.0 (#164), Qwen3.5 35B-A3B: 42.4 (#127)

Long Context benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen3.5 35B-A3B
LMArena Longer Query13151389

Writing & Preference Qwen3.5 35B-A3B leads

Nvidia Llama 3.3 Nemotron Super 49b v1.5: 53.1 (#159), Qwen3.5 35B-A3B: 57.9 (#124)

Writing & Preference benchmarks
BenchmarkNvidia Llama 3.3 Nemotron Super 49b v1.5Qwen3.5 35B-A3B
LMArena Text13381395
LMArena Creative Writing13071346
LMArena Multi-Turn13341390

Frequently asked questions

Is Nvidia Llama 3.3 Nemotron Super 49b v1.5 better than Qwen3.5 35B-A3B?

Qwen3.5 35B-A3B is the stronger model overall, scoring 42.0 to 40.3 on the Noometry Index. Nvidia Llama 3.3 Nemotron Super 49b v1.5 costs 1.7× less per token, which makes it the better buy when Qwen3.5 35B-A3B's lead doesn't matter for your workload.

Which is cheaper, Nvidia Llama 3.3 Nemotron Super 49b v1.5 or Qwen3.5 35B-A3B?

Nvidia Llama 3.3 Nemotron Super 49b v1.5 is cheaper. It lists at $0.40 per million input tokens and $0.40 per million output tokens; Qwen3.5 35B-A3B lists at $0.25 and $2.

Is Nvidia Llama 3.3 Nemotron Super 49b v1.5 or Qwen3.5 35B-A3B better for coding?

Nvidia Llama 3.3 Nemotron Super 49b v1.5 scores higher on coding benchmarks: 39.8 versus 33.8 in the Noometry coding category.

Which has the bigger context window?

Qwen3.5 35B-A3B does, with 262K tokens against 131K.

How many benchmarks do Nvidia Llama 3.3 Nemotron Super 49b v1.5 and Qwen3.5 35B-A3B share?

12 benchmarks have published results for both models. Nvidia Llama 3.3 Nemotron Super 49b v1.5 has 12 scored results on Noometry and Qwen3.5 35B-A3B has 28.

Related comparisons

Go deeper